Skip to content

Is Reinforcement Learning More Difficult Than Bandits? A Near-optimal Algorithm Escaping the Curse of Horizon.

Zihan Zhang, Xiangyang Ji, Simon S. Du

VenueA*COLT
Year2021
ProceedingsCOLT

Browse the full COLT paper archive.