Improving Stochastic Policy Gradients in Continuous Control with Deep Reinforcement Learning using the Beta Distribution.
Po-Wei Chou, Daniel Maturana, Sebastian A. Scherer
Browse the full ICML paper archive.
Po-Wei Chou, Daniel Maturana, Sebastian A. Scherer
Browse the full ICML paper archive.