Skip to content

Maxmin Q-learning: Controlling the Estimation Bias of Q-learning.

Qingfeng Lan, Yangchen Pan, Alona Fyshe, Martha White

VenueA*ICLR
Year2020
ProceedingsICLR

Browse the full ICLR paper archive.