Skip to content

Doubly Robust Off-policy Value Evaluation for Reinforcement Learning.

Nan Jiang, Lihong Li

VenueA*ICML
Year2016
ProceedingsICML

Browse the full ICML paper archive.