← Back to all articles
arXiv cs.CLSeptember 11, 2026

Inverse Turing Bench: Evaluating Language Models as Judges of Human vs. AI Dialogue

Excerpt

arXiv:2606.21844v2 Announce Type: replace Abstract: As AI systems integrate into online spaces, differentiating them from humans in conversations is increasingly important. We present Inverse Turing Bench, a benchmark that evaluates LLMs and other models on their ability to differentiate humans and AI in multi-turn text. The benchmark provides a collection of paired dialogue transcripts, wherein one dialogue is between two humans and the other is between a human and an AI. The task is to correct