Skip to content

Why is Posterior Sampling Better than Optimism for Reinforcement Learning?

Ian Osband, Benjamin Van Roy

VenueA*ICML
Year2017
ProceedingsICML

Browse the full ICML paper archive.