Skip to content

Show, Think, and Tell: Thought-Augmented Fine-Tuning of Large Language Models for Video Captioning.

Byoungjip Kim, Dasol Hwang, Sungjun Cho, Youngsoo Jang, Honglak Lee, Moontae Lee

VenueA*CVPR
Year2024
ProceedingsCVPR Workshops

Browse the full CVPR paper archive.