Skip to content

QCaption: Video Captioning and Q&A through Fusion of Large Multimodal Models.

Jiale Wang, Gee Wah Ng, Lee Onn Mak, Randall Cher, Ng Ding Hei Ryan, Davis Wang

VenueCFUSION
Year2024
ProceedingsFUSION

Browse the full FUSION paper archive.