arXiv cs.AIOctober 7, 2026
Benchmarking Psychological Dynamics in Generative Agents
Excerpt
arXiv:2610.04246v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed to simulate human behavior, acting as computational replicas of human subjects. Yet the lived psychological experience of humans is difficult to benchmark, particularly as it unfolds over time. We introduce a psychometric benchmark for computational replicas: personas that carry a fixed identity through an evolving sequence of events. Built entirely from published norms and meta-analytic effe