Skip to content

Improving Stochastic Policy Gradients in Continuous Control with Deep Reinforcement Learning using the Beta Distribution.

Po-Wei Chou, Daniel Maturana, Sebastian A. Scherer

VenueA*ICML
Year2017
ProceedingsICML

Browse the full ICML paper archive.