← Back to all articles
arXiv cs.LGOctober 7, 2026

Stateless Language Agents: Scaling Long-Horizon Automated Research

Excerpt

arXiv:2610.07625v1 Announce Type: new Abstract: Automated research systems increasingly run LLM agents over long horizons, but more inference does not by itself produce more progress: agents replay growing histories, duplicate one another's work, or stop experimenting while token consumption continues. Yet most evaluations use short budgets or benchmarks that saturate early, leaving these failure modes untested. We trace these failures to two choices: where research state lives and who decides w