Collaborative Static and Dynamic Vision-Language Streams for Spatio-Temporal Video Grounding.
Zihang Lin, Chaolei Tan, Jian-Fang Hu, Zhi Jin, Tiancai Ye, Wei-Shi Zheng
Browse the full CVPR paper archive.
Zihang Lin, Chaolei Tan, Jian-Fang Hu, Zhi Jin, Tiancai Ye, Wei-Shi Zheng
Browse the full CVPR paper archive.