If CLIP Could Talk: Understanding Vision-Language Model Representations Through Their Preferred Concept Descriptions.
Reza Esfandiarpoor, Cristina Menghini, Stephen H. Bach
Browse the full EMNLP paper archive.
Reza Esfandiarpoor, Cristina Menghini, Stephen H. Bach
Browse the full EMNLP paper archive.