Provably efficient reinforcement learning with linear function approximation.
Chi Jin, Zhuoran Yang, Zhaoran Wang, Michael I. Jordan
Browse the full COLT paper archive.
Chi Jin, Zhuoran Yang, Zhaoran Wang, Michael I. Jordan
Browse the full COLT paper archive.