arXiv cs.AIOctober 7, 2026
RTSGameBench: An RTS Benchmark for Strategic Reasoning by Vision-Language Models
Excerpt
arXiv:2606.18950v3 Announce Type: replace Abstract: Modern Vision-Language Models (VLMs) often struggle with strategic reasoning, i.e., anticipating and influencing other agents' actions, under uncertainty in competitive and cooperative settings. Real-time strategy (RTS) games can be a natural testbed for diagnosing this limitation, as they demand coordination with allies, adaptation to opponents' strategy, and long-horizon planning under partial observability. However, existing RTS benchmarks o