papersTODAY 04:00 UTC
SURE-Voice: Training-Free Front End Filters Speech Evidence for Speech LLMs
A new arXiv paper introduces SURE-Voice, a training-free front-end component that estimates whether audio actually contains usable speech evidence before a speech language model generates output. The work frames the problem as pre-generation support estimation, targeting cases where speech LLMs produce plausible but unsupported responses from audio lacking meaningful speech.