Skip to content

Reward Estimation for Variance Reduction in Deep Reinforcement Learning.

Joshua Romoff, Alexandre Pich, Peter Henderson, Vincent Franois-Lavet, Joelle Pineau

VenueA*ICLR
Year2018
ProceedingsICLR (Workshop)

Browse the full ICLR paper archive.