VLCAP: Vision-Language with Contrastive Learning for Coherent Video Paragraph Captioning.
Kashu Yamazaki, Sang Truong, Viet-Khoa Vo-Ho, Michael Kidd, Chase Rainwater, Khoa Luu, Ngan Le
Browse the full ICIP paper archive.
Kashu Yamazaki, Sang Truong, Viet-Khoa Vo-Ho, Michael Kidd, Chase Rainwater, Khoa Luu, Ngan Le
Browse the full ICIP paper archive.