Not All Frames Are Equal: Weakly-Supervised Video Grounding With Contextual Similarity and Visual Clustering Losses.
Jing Shi, Jia Xu, Boqing Gong, Chenliang Xu
Browse the full CVPR paper archive.
Jing Shi, Jia Xu, Boqing Gong, Chenliang Xu
Browse the full CVPR paper archive.