Variation-resistant Q-learning: Controlling and Utilizing Estimation Bias in Reinforcement Learning for Better Performance.
Andreas Pentaliotis, Marco A. Wiering
Browse the full ICAART paper archive.
Andreas Pentaliotis, Marco A. Wiering
Browse the full ICAART paper archive.