Skip to content

MSRS: Training Multimodal Speech Recognition Models from Scratch with Sparse Mask Optimization.

Adriana Fernandez-Lopez, Honglie Chen, Pingchuan Ma, Lu Yin, Qiao Xiao, Stavros Petridis, Shiwei Liu, Maja Pantic

Year2024
ProceedingsINTERSPEECH

Browse the full Interspeech paper archive.