Near-Optimal Model-Free Reinforcement Learning in Non-Stationary Episodic MDPs.
Weichao Mao, Kaiqing Zhang, Ruihao Zhu, David Simchi-Levi, Tamer Basar
Browse the full ICML paper archive.
Weichao Mao, Kaiqing Zhang, Ruihao Zhu, David Simchi-Levi, Tamer Basar
Browse the full ICML paper archive.