Skip to content

VicTR: Video-conditioned Text Representations for Activity Recognition.

Kumara Kahatapitiya, Anurag Arnab, Arsha Nagrani, Michael S. Ryoo

VenueA*CVPR
Year2024
ProceedingsCVPR

Browse the full CVPR paper archive.