Skip to content

STVGBert: A Visual-linguistic Transformer based Framework for Spatio-temporal Video Grounding.

Rui Su, Qian Yu, Dong Xu

VenueA*ICCV
Year2021
ProceedingsICCV

Browse the full ICCV paper archive.