Skip to content

Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor.

Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel, Sergey Levine

VenueA*ICML
Year2018
ProceedingsICML

Browse the full ICML paper archive.