Skip to content

Tighter Problem-Dependent Regret Bounds in Reinforcement Learning without Domain Knowledge using Value Function Bounds.

Andrea Zanette, Emma Brunskill

VenueA*ICML
Year2019
ProceedingsICML

Browse the full ICML paper archive.