Skip to content

Learning Curve Bounds for a Markov Decision Process with Undiscounted Rewards.

Lawrence K. Saul, Satinder P. Singh

VenueA*COLT
Year1996
ProceedingsCOLT

Browse the full COLT paper archive.