← Back to all articles
arXiv cs.CLSeptember 24, 2026

PASTABench: Proactive Assessment of Sequential Trajectories for Agent Safety

Excerpt

arXiv:2609.28197v1 Announce Type: cross Abstract: As Large Language Models (LLMs) evolve into autonomous agents that alter real-world states, ensuring operational safety across multi-step workflows has become a critical challenge. While recent work has moved beyond single-turn evaluation toward multi-turn paradigms, key limitations persist: step-level methods treat actions in isolation, missing how risks accumulate, while trajectory-level evaluations operate post-hoc, offering no opportunity for