Pessimistic Q-Learning for Offline Reinforcement Learning: Towards Optimal Sample Complexity.
Laixi Shi, Gen Li, Yuting Wei, Yuxin Chen, Yuejie Chi
Browse the full ICML paper archive.
Laixi Shi, Gen Li, Yuting Wei, Yuxin Chen, Yuejie Chi
Browse the full ICML paper archive.