arXiv cs.AIOctober 2, 2026
The Hitchhikers Guide to Rubric Quality Understanding and Enrichment
Excerpt
arXiv:2604.01375v3 Announce Type: replace Abstract: Rubrics distill notions of expert quality and measure agent performance. However, the quality of rubrics themselves have not been systematically measured and are often left to downstream performance.We import apparatuses from measurement theory built for exactly this: quantitative signals based on the rubric's content, and introduce the RubrIc-Failure Taxonomy (RIFT), of nine possible ways a rubric fails, organized under reliability and content