Skip to content

Text-Infused Audio-Visual Video Parsing with Semantic-Aware Multimodal Contrastive Learning.

Pengcheng Zhao, Yanxiang Chen, Dan Guo, Yuanzhi Yao

Year2025
ProceedingsICASSP

Browse the full ICASSP paper archive.