Are Vision-Language Transformers Learning Multimodal Representations? A Probing Perspective.
Emmanuelle Salin, Badreddine Farah, Stphane Ayache, Benot Favre
Browse the full AAAI paper archive.
Emmanuelle Salin, Badreddine Farah, Stphane Ayache, Benot Favre
Browse the full AAAI paper archive.