← Back to all articles
arXiv cs.AIOctober 7, 2026

Benchmarking Psychological Dynamics in Generative Agents

Excerpt

arXiv:2610.04246v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed to simulate human behavior, acting as computational replicas of human subjects. Yet the lived psychological experience of humans is difficult to benchmark, particularly as it unfolds over time. We introduce a psychometric benchmark for computational replicas: personas that carry a fixed identity through an evolving sequence of events. Built entirely from published norms and meta-analytic effe