Skip to content

Diffusion Bridge: Leveraging Diffusion Model to Reduce the Modality Gap Between Text and Vision for Zero-Shot Image Captioning.

Jeong Ryong Lee, Yejee Shin, Geonhui Son, Dosik Hwang

VenueA*CVPR
Year2025
ProceedingsCVPR

Browse the full CVPR paper archive.