Skip to content

VL-CLIP: Enhancing Multimodal Recommendations via Visual Grounding and LLM-Augmented CLIP Embeddings.

Ramin Giahi, Kehui Yao, Sriram Kollipara, Kai Zhao, Vahid Mirjalili, Jianpeng Xu, Topojoy Biswas, Evren Krpeoglu, Kannan Achan

VenueARecSys
Year2025
ProceedingsRecSys

Browse the full RecSys paper archive.