Skip to content

Convergence of Policy Gradient for Entropy Regularized MDPs with Neural Network Approximation in the Mean-Field Regime.

James-Michael Leahy, Bekzhan Kerimkulov, David Siska, Lukasz Szpruch

VenueA*ICML
Year2022
ProceedingsICML

Browse the full ICML paper archive.