Can Vision-Language Models Understand and Interpret Dynamic Gestures from Pedestrians? Pilot Datasets and Exploration Towards Instructive Nonverbal Commands for Cooperative Autonomous Vehicles.
Tonko E. W. Bossen, Andreas Mgelmose, Ross Greer
Browse the full CVPR paper archive.