Skip to content

Knowledge Completes the Vision: A Multimodal Entity-aware Retrieval-Augmented Generation Framework for News Image Captioning.

Xiaoxing You, Qiang Huang, Lingyu Li, Chi Zhang, Xiaopeng Liu, Min Zhang, Jun Yu

VenueA*AAAI
Year2026
ProceedingsAAAI

Browse the full AAAI paper archive.