Skip to content

Speech ReaLLM - Real-time Speech Recognition with Multimodal Language Models by Teaching the Flow of Time.

Frank Seide, Yangyang Shi, Morrie Doulaty, Yashesh Gaur, Junteng Jia, Chunyang Wu

Year2024
ProceedingsINTERSPEECH

Browse the full Interspeech paper archive.