Skip to content

VLASCD: A Visual Language Action Model for Simultaneous Chatting and Decision Making.

Zuojin Tang, Bin Hu, Chenyang Zhao, De Ma, Gang Pan, Bin Liu

VenueA*EMNLP
Year2025
ProceedingsEMNLP

Browse the full EMNLP paper archive.