Skip to content

LAVCap: LLM-based Audio-Visual Captioning using Optimal Transport.

Kyeongha Rho, Hyeongkeun Lee, Valentio Iverson, Joon Son Chung

Year2025
ProceedingsICASSP

Browse the full ICASSP paper archive.