Diffusion Bridge: Leveraging Diffusion Model to Reduce the Modality Gap Between Text and Vision for Zero-Shot Image Captioning.
Jeong Ryong Lee, Yejee Shin, Geonhui Son, Dosik Hwang
Browse the full CVPR paper archive.
Jeong Ryong Lee, Yejee Shin, Geonhui Son, Dosik Hwang
Browse the full CVPR paper archive.