Audio LLMs assess reliability of their own transcriptions; arXiv:2609.30625 shows models can recognize when transcription may be unreliable
Read the original at arxiv.org→arXiv:2609.30625v1 Announce Type: new Abstract: Audio large language models allow users to interact with the model through speech. When an input recording is too degraded, the model may misinterpret the user's query...
Original headline: "Audio LLMs Know When They Can't Hear You"
Coverage timeline
- Sep 28, 04:00 UTC arXiv cs.AI lead source Audio LLMs Know When They Can't Hear You