Skip to content

Neural Policy Gradient Methods: Global Optimality and Rates of Convergence.

Lingxiao Wang, Qi Cai, Zhuoran Yang, Zhaoran Wang

VenueA*ICLR
Year2020
ProceedingsICLR

Browse the full ICLR paper archive.