Skip to content

Guided Dialog Policy Learning: Reward Estimation for Multi-Domain Task-Oriented Dialog.

Ryuichi Takanobu, Hanlin Zhu, Minlie Huang

VenueA*EMNLP
Year2019
ProceedingsEMNLP/IJCNLP (1)

Browse the full EMNLP paper archive.