Skip to content

Pessimistic Q-Learning for Offline Reinforcement Learning: Towards Optimal Sample Complexity.

Laixi Shi, Gen Li, Yuting Wei, Yuxin Chen, Yuejie Chi

VenueA*ICML
Year2022
ProceedingsICML

Browse the full ICML paper archive.