Skip to content

MMFT-BERT: Multimodal Fusion Transformer with BERT Encodings for Visual Question Answering.

Aisha Urooj Khan, Amir Mazaheri, Niels da Vitoria Lobo, Mubarak Shah

VenueA*EMNLP
Year2020
ProceedingsEMNLP (Findings)

Browse the full EMNLP paper archive.