arXiv cs.LGOctober 2, 2026
Temporally-Resolved Token Attribution Reveals the Generation Dynamics of Diffusion Language Models
Excerpt
arXiv:2610.01177v1 Announce Type: cross Abstract: This work presents Diffusion Layer Integrated Gradients (DLIG), a token attribution method for diffusion language models (DLMs) that extends Integrated Gradients (IG~\cite{sundararajan2017axiomatic}) to arbitrary layers and denoising steps. DLIG attributes a DLM's progressive commitment to a self-generated or fixed completion for an input prompt. We establish direct correspondences between DLIG and the IG axioms of completeness, implementation in