Can LLMs Deceive CLIP? Benchmarking Adversarial Compositionality of Pre-trained Multimodal Representation via Text Updates.
Jaewoo Ahn, Heeseung Yun, Dayoon Ko, Gunhee Kim
Browse the full ACL paper archive.
Jaewoo Ahn, Heeseung Yun, Dayoon Ko, Gunhee Kim
Browse the full ACL paper archive.