Skip to content

Reinforcement learning for dialog management using least-squares Policy iteration and fast feature selection.

Lihong Li, Jason D. Williams, Suhrid Balakrishnan

Year2009
ProceedingsINTERSPEECH

Browse the full Interspeech paper archive.