Skip to content

DR3: Value-Based Deep Reinforcement Learning Requires Explicit Regularization.

Aviral Kumar, Rishabh Agarwal, Tengyu Ma, Aaron C. Courville, George Tucker, Sergey Levine

VenueA*ICLR
Year2022
ProceedingsICLR

Browse the full ICLR paper archive.