Geometrically-Aware Dual Transformer Encoding Visual and Textual Features for Image Captioning.
Yu-Ling Chang, Hao-Shang Ma, Shiou-Chi Li, Jen-Wei Huang
Browse the full PAKDD paper archive.
Yu-Ling Chang, Hao-Shang Ma, Shiou-Chi Li, Jen-Wei Huang
Browse the full PAKDD paper archive.