Skip to content

Disambiguating Reference in Visually Grounded Dialogues through Joint Modeling of Textual and Multimodal Semantic Structures.

Shun Inadumi, Nobuhiro Ueda, Koichiro Yoshino

VenueA*ACL
Year2025
ProceedingsACL (1)

Browse the full ACL paper archive.