Skip to content

Black-box Off-policy Estimation for Infinite-Horizon Reinforcement Learning.

Ali Mousavi, Lihong Li, Qiang Liu, Denny Zhou

VenueA*ICLR
Year2020
ProceedingsICLR

Browse the full ICLR paper archive.