Skip to content

Asymptotically Efficient Off-Policy Evaluation for Tabular Reinforcement Learning.

Ming Yin, Yu-Xiang Wang

Year2020
ProceedingsAISTATS

Browse the full AISTATS paper archive.