Skip to content

Captions Speak Louder than Images: Generalizing Foundation Models for E-commerce from High-quality Multimodal Instruction Data.

Xinyi Ling, Hanwen Du, Bo Peng, Zhihui Zhu, Xia Ning

VenueBIJCNLP
Year2025
ProceedingsIJCNLP-AACL (long papers)

Browse the full IJCNLP paper archive.