Multi-agent cooperation through learning-aware policy gradients.
Alexander Meulemans, Seijin Kobayashi, Johannes von Oswald, Nino Scherrer, Eric Elmoznino, Blake Aaron Richards, Guillaume Lajoie, Blaise Agera y Arcas, Joo Sacramento
Browse the full ICLR paper archive.