← Back to all articles
arXiv cs.CLSeptember 24, 2026

Text Scores Can Miss Waveform Use: A Qwen2-Audio Quantization Case Study

Excerpt

arXiv:2609.26823v1 Announce Type: cross Abstract: Post-training quantization of speech language models is often summarized with text-output scores and nominal bit widths. Those numbers alone do not establish behavior that depends on information missing from a transcript, or efficiency for a particular runtime. We introduce an evaluation protocol that separately tests lexical output, a transcript-insufficient endpoint, and a measured packed implementation. In a Qwen2-Audio case study, a translati