Skip to content

Optimality and Approximation with Policy Gradient Methods in Markov Decision Processes.

Alekh Agarwal, Sham M. Kakade, Jason D. Lee, Gaurav Mahajan

VenueA*COLT
Year2020
ProceedingsCOLT

Browse the full COLT paper archive.