Skip to content

Act-ChatGPT: Introducing Action Features into Multi-modal Large Language Models for Video Understanding.

Yuto Nakamizo, Keiji Yanai

VenueBICPR
Year2024
ProceedingsICPR (23)

Browse the full ICPR paper archive.