arXiv cs.CLSeptember 11, 2026
Beyond Solver Verdicts: Generative Reward Models for Autoformalization
Excerpt
arXiv:2609.11085v1 Announce Type: cross Abstract: Neurosymbolic systems rely on mathematical solvers to guarantee reasoning correctness, yet solvers are fundamentally blind to whether a formal translation maintains strict reference-equivalence to a designated formalization. We formalize this vulnerability as Verdict-Preserving-Unfaithfulness (VPU): a failure mode where an incorrect encoding executes successfully and matches the expected verdict. We theoretically prove that structural, verdict-on