Reward Estimation for Variance Reduction in Deep Reinforcement Learning.
Joshua Romoff, Peter Henderson, Alexandre Pich, Vincent Franois-Lavet, Joelle Pineau
Browse the full CoRL paper archive.
Joshua Romoff, Peter Henderson, Alexandre Pich, Vincent Franois-Lavet, Joelle Pineau
Browse the full CoRL paper archive.