← Back to all articles
arXiv cs.AIOctober 7, 2026

GNN-CB: A Graph Neural Network Competition Benchmark for Human and LLM Evaluation

Excerpt

arXiv:2610.05387v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated strong performance on coding and reasoning benchmarks; however, their ability to solve graph-structured machine learning problems remains largely unexplored. In particular, no benchmark currently evaluates whether LLMs can autonomously solve end-to-end Graph Neural Network (GNN) coding tasks under realistic competition settings. To address this gap, this paper introduces GNN-CB, the first competition