Nearly Minimax Optimal Reinforcement Learning for Linear Markov Decision Processes.
Jiafan He, Heyang Zhao, Dongruo Zhou, Quanquan Gu
Browse the full ICML paper archive.
Jiafan He, Heyang Zhao, Dongruo Zhou, Quanquan Gu
Browse the full ICML paper archive.