Skip to content

Track the Answer: Extending TextVQA from Image to Video with Spatio-Temporal Clues.

Yan Zhang, Gangyan Zeng, Huawen Shen, Daiqing Wu, Yu Zhou, Can Ma

VenueA*AAAI
Year2025
ProceedingsAAAI

Browse the full AAAI paper archive.