First-Order Regret in Reinforcement Learning with Linear Function Approximation: A Robust Estimation Approach.
Andrew J. Wagenmaker, Yifang Chen, Max Simchowitz, Simon S. Du, Kevin Jamieson
Browse the full ICML paper archive.
Andrew J. Wagenmaker, Yifang Chen, Max Simchowitz, Simon S. Du, Kevin Jamieson
Browse the full ICML paper archive.