Distributed Audio-Visual Parsing Based On Multimodal Transformer and Deep Joint Source Channel Coding.
Penghong Wang, Jiahui Li, Mengyao Ma, Xiaopeng Fan
Browse the full ICASSP paper archive.
Penghong Wang, Jiahui Li, Mengyao Ma, Xiaopeng Fan
Browse the full ICASSP paper archive.