Reinforcement learning for dialog management using least-squares Policy iteration and fast feature selection.
Lihong Li, Jason D. Williams, Suhrid Balakrishnan
Browse the full Interspeech paper archive.
Lihong Li, Jason D. Williams, Suhrid Balakrishnan
Browse the full Interspeech paper archive.