Skip to content

GroundVLP: Harnessing Zero-Shot Visual Grounding from Vision-Language Pre-training and Open-Vocabulary Object Detection.

Haozhan Shen, Tiancheng Zhao, Mingwei Zhu, Jianwei Yin

VenueA*AAAI
Year2024
ProceedingsAAAI

Browse the full AAAI paper archive.