Going Beyond Nouns With Vision & Language Models Using Synthetic Data.
Paola Cascante-Bonilla, Khaled Shehada, James Seale Smith, Sivan Doveh, Donghyun Kim, Rameswar Panda, Gl Varol, Aude Oliva, Vicente Ordonez, Rogrio Feris, Leonid Karlinsky
Browse the full ICCV paper archive.