Skip to content

Learning Near-Optimal Policies with Bellman-Residual Minimization Based Fitted Policy Iteration and a Single Sample Path.

Andrs Antos, Csaba Szepesvri, Rmi Munos

VenueA*COLT
Year2006
ProceedingsCOLT

Browse the full COLT paper archive.