Symmetric (Optimistic) Natural Policy Gradient for Multi-Agent Learning with Parameter Convergence.
Sarath Pattathil, Kaiqing Zhang, Asuman E. Ozdaglar
Browse the full AISTATS paper archive.
Sarath Pattathil, Kaiqing Zhang, Asuman E. Ozdaglar
Browse the full AISTATS paper archive.