Trajectory-Based Off-Policy Deep Reinforcement Learning.
Andreas Doerr, Michael Volpp, Marc Toussaint, Sebastian Trimpe, Christian Daniel
Browse the full ICML paper archive.
Andreas Doerr, Michael Volpp, Marc Toussaint, Sebastian Trimpe, Christian Daniel
Browse the full ICML paper archive.