DR3: Value-Based Deep Reinforcement Learning Requires Explicit Regularization.
Aviral Kumar, Rishabh Agarwal, Tengyu Ma, Aaron C. Courville, George Tucker, Sergey Levine
Browse the full ICLR paper archive.
Aviral Kumar, Rishabh Agarwal, Tengyu Ma, Aaron C. Courville, George Tucker, Sergey Levine
Browse the full ICLR paper archive.