Root-n-Regret for Learning in Markov Decision Processes with Function Approximation and Low Bellman Rank.
Kefan Dong, Jian Peng, Yining Wang, Yuan Zhou
Browse the full COLT paper archive.
Kefan Dong, Jian Peng, Yining Wang, Yuan Zhou
Browse the full COLT paper archive.