Skip to content

TMT: A Transformer-Based Modal Translator for Improving Multimodal Sequence Representations in Audio Visual Scene-Aware Dialog.

Wubo Li, Dongwei Jiang, Wei Zou, Xiangang Li

Year2020
ProceedingsINTERSPEECH

Browse the full Interspeech paper archive.