Neural Rewards Regression for near-optimal policy identification in Markovian and partial observable environments.
Daniel Schneega, Steffen Udluft, Thomas Martinetz
Browse the full ESANN paper archive.
Daniel Schneega, Steffen Udluft, Thomas Martinetz
Browse the full ESANN paper archive.