Skip to content

MMSpeech: Multi-modal Multi-task Encoder-Decoder Pre-training for speech recognition.

Xiaohuan Zhou, Jiaming Wang, Zeyu Cui, Shiliang Zhang, Zhijie Yan, Jingren Zhou, Chang Zhou

Year2023
ProceedingsINTERSPEECH

Browse the full Interspeech paper archive.