← Back to all articles
arXiv cs.AIOctober 7, 2026

MASBench: Benchmarking LLM-based Multi-Agent Collaboration under Partial Observability

Excerpt

arXiv:2610.04672v1 Announce Type: new Abstract: Large language models (LLMs) have progressively evolved into the core of autonomous agents. Building on this progress, LLM-based multi-agent systems (MAS) coordinate multiple agents into a synergistic team to accomplish complex tasks that exceed the capabilities of individual agents. The effectiveness of such systems depends not only on the agents themselves, but also on how collaboration mechanisms are designed and organized. Note that real-world