Skip to content

Problem Dependent Reinforcement Learning Bounds Which Can Identify Bandit Structure in MDPs.

Andrea Zanette, Emma Brunskill

VenueA*ICML
Year2018
ProceedingsICML

Browse the full ICML paper archive.