Skip to content

Learning to See through Sound: From VggCaps to Multi2Cap for Richer Automated Audio Captioning.

Sangyeon Cho, Mingi Kim, Jinkwon Hwang, Jaehoon Go, Minuk Ma, Sunjae Yoon, Junyeong Kim

VenueA*EMNLP
Year2025
ProceedingsEMNLP

Browse the full EMNLP paper archive.