KG-ViP: Bridging Knowledge Grounding and Visual Perception in Multi-modal LLMs for Visual Question Answering.
Zhiyang Li, Ao Ke, Yukun Cao, Xike Xie
Browse the full ACL paper archive.
Zhiyang Li, Ao Ke, Yukun Cao, Xike Xie
Browse the full ACL paper archive.