Multimodal Speech Recognition for Language-Guided Embodied Agents.
Allen Chang, Xiaoyuan Zhu, Aarav Monga, Seoho Ahn, Tejas Srinivasan, Jesse Thomason
Browse the full Interspeech paper archive.
Allen Chang, Xiaoyuan Zhu, Aarav Monga, Seoho Ahn, Tejas Srinivasan, Jesse Thomason
Browse the full Interspeech paper archive.