Skip to content

Sample Efficient On-Line Learning of Optimal Dialogue Policies with Kalman Temporal Differences.

Olivier Pietquin, Matthieu Geist, Senthilkumar Chandramohan

VenueA*IJCAI
Year2011
ProceedingsIJCAI

Browse the full IJCAI paper archive.