Skip to content

Double Reinforcement Learning for Efficient and Robust Off-Policy Evaluation.

Nathan Kallus, Masatoshi Uehara

VenueA*ICML
Year2020
ProceedingsICML

Browse the full ICML paper archive.