Skip to content

On-line Active Reward Learning for Policy Optimisation in Spoken Dialogue Systems.

Pei-Hao Su, Milica Gasic, Nikola Mrksic, Lina Maria Rojas-Barahona, Stefan Ultes, David Vandyke, Tsung-Hsien Wen, Steve J. Young

VenueA*ACL
Year2016
ProceedingsACL (1)

Browse the full ACL paper archive.