On Reward-Free RL with Kernel and Neural Function Approximations: Single-Agent MDP and Markov Game.
Shuang Qiu, Jieping Ye, Zhaoran Wang, Zhuoran Yang
Browse the full ICML paper archive.
Shuang Qiu, Jieping Ye, Zhaoran Wang, Zhuoran Yang
Browse the full ICML paper archive.