Skip to content

ViCocktail: Automated Multi-Modal Data Collection for Vietnamese Audio-Visual Speech Recognition.

Thai-Binh Nguyen, Thi Van Nguyen, Quoc Truong Do, Chi Mai Luong

Year2025
ProceedingsINTERSPEECH

Browse the full Interspeech paper archive.