arXiv NLP
techCenter
Listen to the Latents: Self-Correcting Speech Recognition in Large Audio Language Models Through Hidden-State Interactionstranslating…
arXiv:2609.02940v1 Announce Type: new
Abstract: Recent automatic speech recognition (ASR) systems increasingly integrate large language models (LLMs) to leverage their semantic knowledge, either externally through logit fusion or internally through warm initialization. However…
Keywords#Through#Language Models#Speech Recognition#Large#Audio
Comments
Sign in to join the discussion.
Loading comments…