Locate Then Generate: Bridging Vision and Language with Bounding Box for Scene-Text VQA.
Yongxin Zhu, Zhen Liu, Yukang Liang, Xin Li, Hao Liu, Changcun Bao, Linli Xu
Browse the full AAAI paper archive.
Yongxin Zhu, Zhen Liu, Yukang Liang, Xin Li, Hao Liu, Changcun Bao, Linli Xu
Browse the full AAAI paper archive.