Audio LLMs Know When They Can't Hear You

cs.AI updates on arXiv.org · 2h ago
Research Papers

arXiv:2609.30625v1 Announce Type: new Abstract: Audio large language models allow users to interact with the model through speech. When an input recording is too degraded, the model may misinterpret the user's query and respond based on an incorrect transcription. In this paper, we study model-conditional transcription reliability: whether an Audio LLM can recognize when its own transcription is…

Read original article on cs.AI updates on arXiv.org →