arXiv cs.AIOctober 7, 2026
Measuring Intelligence Beyond Human Scale
Excerpt
arXiv:2607.07040v2 Announce Type: replace Abstract: How can we measure intelligence beyond human capability? Human-authored benchmarks saturate, and above human capability, examiners may not know which tasks are both hard and verifiable. We argue that this difficulty is inherent to absolute-scale evaluation and propose a new paradigm based on relative measurement in which models generate public challenges that separate other systems. Aggregating these outcomes yields an adversarial psychometric