Bridging the Visual Gap: Fine-Tuning Multimodal Models with Knowledge-Adapted Captions.
Moran Yanuka, Assaf Ben-Kish, Yonatan Bitton, Idan Szpektor, Raja Giryes
Browse the full NAACL paper archive.
Moran Yanuka, Assaf Ben-Kish, Yonatan Bitton, Idan Szpektor, Raja Giryes
Browse the full NAACL paper archive.