Align-KD: Distilling Cross-Modal Alignment Knowledge for Mobile Vision-Language Large Model Enhancement.
Qianhan Feng, Wenshuo Li, Tong Lin, Xinghao Chen
Browse the full CVPR paper archive.
Qianhan Feng, Wenshuo Li, Tong Lin, Xinghao Chen
Browse the full CVPR paper archive.