arXiv cs.AIOctober 7, 2026
Organising Trajectory Evidence for Language-Model Agent Assurance: Fragments, Methods, and the Residual
Excerpt
arXiv:2610.04710v1 Announce Type: cross Abstract: Methods for assessing language-model agents include rule checkers over logs, analyses of skill coverage and composition, support checkers, prefix monitors, execution gates, and rare-event estimators. Each observes a different part of a run and makes a claim of a different strength, and no common account says how these claims combine or what they leave unchecked. We give one, built from a logic, information fragments, and an assurance ledger. Requ