Skip to content

Spatiotemporal-Aware Visual Captioning using Vision-Language Pre-Training Model.

Shuai Wu, Weidong Yang, Shuyan Wu

Year2025
ProceedingsICASSP

Browse the full ICASSP paper archive.