← Back to all articles
arXiv cs.CLSeptember 22, 2026

LLaDA-PRM: A Bidirectional Step-Level Reasoning Evaluator

Excerpt

arXiv:2609.22700v1 Announce Type: new Abstract: Step-level reasoning evaluators are commonly based on autoregressive language models, whose causal attention restricts each step representation to the problem, previous steps, and the current step. Yet, when the complete solution is available, the validity of an earlier step may become clearer only through its downstream consequences. We validate this hypothesis through a controlled 54-run comparison of causal and bidirectional LLaDA evaluators at