QCaption: Video Captioning and Q&A through Fusion of Large Multimodal Models.
Jiale Wang, Gee Wah Ng, Lee Onn Mak, Randall Cher, Ng Ding Hei Ryan, Davis Wang
Browse the full FUSION paper archive.
Jiale Wang, Gee Wah Ng, Lee Onn Mak, Randall Cher, Ng Ding Hei Ryan, Davis Wang
Browse the full FUSION paper archive.