Skip to content

Doubling down: sparse grounding with an additional, almost-matching caption for detection-oriented multimodal pretraining.

Giacomo Nebbia, Adriana Kovashka

VenueA*CVPR
Year2022
ProceedingsCVPR Workshops

Browse the full CVPR paper archive.