Skip to content

MM-CARP: Multimodal Model with Cross-Modal Retrieval-Augmented and Visual Region Perception.

Junhao Guo, Chenhan Fu, Guoming Wang, Rongxing Lu, Dong Chen, Siliang Tang

VenueBMMM
Year2025
ProceedingsMMM (2)

Browse the full MMM paper archive.