Skip to content

Incorporating Scene Graphs into Pre-trained Vision-Language Models for Multimodal Open-vocabulary Action Recognition.

Chao Wei, Zhidong Deng

VenueA*ICRA
Year2024
ProceedingsICRA

Browse the full ICRA paper archive.