← Back to all articles
arXiv cs.AIOctober 7, 2026

Measuring Intelligence Beyond Human Scale

Excerpt

arXiv:2607.07040v2 Announce Type: replace Abstract: How can we measure intelligence beyond human capability? Human-authored benchmarks saturate, and above human capability, examiners may not know which tasks are both hard and verifiable. We argue that this difficulty is inherent to absolute-scale evaluation and propose a new paradigm based on relative measurement in which models generate public challenges that separate other systems. Aggregating these outcomes yields an adversarial psychometric