A Temporal Difference GNG-Based Algorithm That Can Learn to Control in Reinforcement Learning Environments.
Davi C. de L. Vieira, Paulo J. L. Adeodato, Paulo M. Goncalves Junior
Browse the full ICMLA paper archive.
Davi C. de L. Vieira, Paulo J. L. Adeodato, Paulo M. Goncalves Junior
Browse the full ICMLA paper archive.