← Back to all articles
arXiv cs.AIAugust 18, 2026

Reconstruction: A Blind Benchmark for Recovering Research Ideas from Pre-Publication Bibliographies

Excerpt

arXiv:2608.16645v1 Announce Type: new Abstract: Can a language model recover the true research idea of a published paper when given only that paper's pre-publication bibliography? We introduce Reconstruction, a blind idea-recovery benchmark that withholds the seed paper and all contemporaneous or future literature, and asks models to propose hypotheses that an independent large language model judge matches against the held-out ground-truth idea. A strict anti-leakage protocol-temporal citation c