Actual Return Reinforcement Learning versus Temporal Differences: Some Theoretical and Experimental Results.
Mark D. Pendrith, Malcolm R. K. Ryan
Browse the full ICML paper archive.
Mark D. Pendrith, Malcolm R. K. Ryan
Browse the full ICML paper archive.