Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor.
Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel, Sergey Levine
Browse the full ICML paper archive.
Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel, Sergey Levine
Browse the full ICML paper archive.