← Back to all articles
arXiv cs.CLSeptember 22, 2026

Euston: Training Away Mathematical Sycophancy Without Losing the Mathematics

Excerpt

arXiv:2609.23205v1 Announce Type: new Abstract: Reasoning language models are trained to produce solutions, not to refuse them, and this bias persists when the problem they are handed is false. Asked to prove a corrupted theorem, a strong model will typically comply and produce a confident derivation of something untrue. We present Euston, an 8B mathematical claim-verification model trained to resist exactly this. Training data were generated with GraphSynth, a probabilistic factor-graph generat