← Back to all articles
arXiv cs.CLSeptember 10, 2026

$S^3$-Bench: Evaluating Speech Interaction Models as Scientific Voice Assistants

Excerpt

arXiv:2609.09852v1 Announce Type: new Abstract: The advance of multimodal large language models (MLLMs) has fundamentally reshaped the paradigm of human-computer interaction, especially speech interaction models capable of seamless conversations. Despite remarkable performance as general voice assistants, their performance in specialized domains remains underexplored, particularly in scientific areas. Scientific interactions introduce formidable challenges, involving rare technical terminology,