Value learning from trajectory optimization and Sobolev descent: A step toward reinforcement learning with superlinear convergence properties.
Amit Parag, Sbastien Kleff, Lo Saci, Nicolas Mansard, Olivier Stasse
Browse the full ICRA paper archive.
Amit Parag, Sbastien Kleff, Lo Saci, Nicolas Mansard, Olivier Stasse
Browse the full ICRA paper archive.