Captions Speak Louder than Images: Generalizing Foundation Models for E-commerce from High-quality Multimodal Instruction Data.
Xinyi Ling, Hanwen Du, Bo Peng, Zhihui Zhu, Xia Ning
Browse the full IJCNLP paper archive.
Xinyi Ling, Hanwen Du, Bo Peng, Zhihui Zhu, Xia Ning
Browse the full IJCNLP paper archive.