← Back to all articles
arXiv cs.LGOctober 2, 2026

LAST: Looped Audio Spectrogram Transformer

Excerpt

arXiv:2610.01926v1 Announce Type: cross Abstract: Increasing depth of transformer models improves recognition, but it comes at a substantial cost. Each additional layer requires more parameters, which makes the process computationally inefficient. We ask whether additional processing can focus on integrating features already computed. Looped Audio Spectrogram Transformer (LAST) first processes all tokens, then reuses the same blocks to refine only the class token over fixed audio features, there