Skip to content

CLIP-ViP: Adapting Pre-trained Image-Text Model to Video-Language Alignment.

Hongwei Xue, Yuchong Sun, Bei Liu, Jianlong Fu, Ruihua Song, Houqiang Li, Jiebo Luo

VenueA*ICLR
Year2023
ProceedingsICLR

Browse the full ICLR paper archive.