Skip to content

Optimal and Adaptive Off-policy Evaluation in Contextual Bandits.

Yu-Xiang Wang, Alekh Agarwal, Miroslav Dudk

VenueA*ICML
Year2017
ProceedingsICML

Browse the full ICML paper archive.