Visual-Text Cross Alignment: Refining the Similarity Score in Vision-Language Models.
Jinhao Li, Haopeng Li, Sarah Monazam Erfani, Lei Feng, James Bailey, Feng Liu
Browse the full ICML paper archive.
Jinhao Li, Haopeng Li, Sarah Monazam Erfani, Lei Feng, James Bailey, Feng Liu
Browse the full ICML paper archive.