Skip to content

Minimax Weight and Q-Function Learning for Off-Policy Evaluation.

Masatoshi Uehara, Jiawei Huang, Nan Jiang

VenueA*ICML
Year2020
ProceedingsICML

Browse the full ICML paper archive.