Point-wise Q-value maximization for converging Q-learning in continuous state-spaces.
Philipp Wissmann, Daniel Hein, Steffen Udluft, Thomas A. Runkler
Browse the full ESANN paper archive.
Philipp Wissmann, Daniel Hein, Steffen Udluft, Thomas A. Runkler
Browse the full ESANN paper archive.