Fixed-Horizon Temporal Difference Methods for Stable Reinforcement Learning.
Kristopher De Asis, Alan Chan, Silviu Pitis, Richard S. Sutton, Daniel Graves
Browse the full AAAI paper archive.
Kristopher De Asis, Alan Chan, Silviu Pitis, Richard S. Sutton, Daniel Graves
Browse the full AAAI paper archive.