Skip to content

If CLIP Could Talk: Understanding Vision-Language Model Representations Through Their Preferred Concept Descriptions.

Reza Esfandiarpoor, Cristina Menghini, Stephen H. Bach

VenueA*EMNLP
Year2024
ProceedingsEMNLP

Browse the full EMNLP paper archive.