arXiv cs.AIOctober 7, 2026
Level-of-Token Diffusion
Excerpt
arXiv:2610.05816v1 Announce Type: cross Abstract: Image and video diffusion models allocate equal computation to every region, even when the intended scene calls for varying levels of detail. The spatial distribution of detail can often be anticipated before generation, indicating where computation can be reduced. We introduce Level-of-Token (LoT) Diffusion, a framework that turns this knowledge into an explicit multiresolution token layout (Level-of-Token layout) for adaptive and efficient gene