Skip to content

Large Language Models Know What is Key Visual Entity: An LLM-assisted Multimodal Retrieval for VQA.

Pu Jian, Donglei Yu, Jiajun Zhang

VenueA*EMNLP
Year2024
ProceedingsEMNLP

Browse the full EMNLP paper archive.