Skip to content

Root-n-Regret for Learning in Markov Decision Processes with Function Approximation and Low Bellman Rank.

Kefan Dong, Jian Peng, Yining Wang, Yuan Zhou

VenueA*COLT
Year2020
ProceedingsCOLT

Browse the full COLT paper archive.