Skip to content

Reinforcement Learning for Bandit Neural Machine Translation with Simulated Human Feedback.

Khanh Nguyen, Hal Daum III, Jordan L. Boyd-Graber

VenueA*EMNLP
Year2017
ProceedingsEMNLP

Browse the full EMNLP paper archive.