MMSpeech: Multi-modal Multi-task Encoder-Decoder Pre-training for speech recognition.
Xiaohuan Zhou, Jiaming Wang, Zeyu Cui, Shiliang Zhang, Zhijie Yan, Jingren Zhou, Chang Zhou
Browse the full Interspeech paper archive.
Xiaohuan Zhou, Jiaming Wang, Zeyu Cui, Shiliang Zhang, Zhijie Yan, Jingren Zhou, Chang Zhou
Browse the full Interspeech paper archive.