Video-Text Prompting for Weakly Supervised Spatio-Temporal Video Grounding.
Heng Zhao, Yinjie Zhao, Bihan Wen, Yew-Soon Ong, Joey Zhou
Browse the full EMNLP paper archive.
Heng Zhao, Yinjie Zhao, Bihan Wen, Yew-Soon Ong, Joey Zhou
Browse the full EMNLP paper archive.