← Back to all articles
arXiv cs.AIOctober 7, 2026

RTSGameBench: An RTS Benchmark for Strategic Reasoning by Vision-Language Models

Excerpt

arXiv:2606.18950v3 Announce Type: replace Abstract: Modern Vision-Language Models (VLMs) often struggle with strategic reasoning, i.e., anticipating and influencing other agents' actions, under uncertainty in competitive and cooperative settings. Real-time strategy (RTS) games can be a natural testbed for diagnosing this limitation, as they demand coordination with allies, adaptation to opponents' strategy, and long-horizon planning under partial observability. However, existing RTS benchmarks o