Skip to content

Bridging Audio and Vision: Zero-Shot Audiovisual Segmentation by Connecting Pretrained Models.

Seung-jae Lee, Paul Hongsuck Seo

Year2025
ProceedingsINTERSPEECH

Browse the full Interspeech paper archive.