← Back to all articles
arXiv cs.AIAugust 17, 2026

Wrong but Useful: Trajectory Value Beyond Answer Correctness in Multi-Agent Messages

Excerpt

arXiv:2608.14375v1 Announce Type: new Abstract: Multi-agent reasoning systems often use agreement, confidence, or automated scores to decide which messages should shape a final answer. Such filtering assumes that a message likely to be correct is also worth keeping. Yet a wrong answer can contain a useful decomposition, constraint, or scientific principle. We test this distinction with Diverse Hypothesis Deliberation (DHD), a controlled measurement protocol that caches five independently generat