Skip to content

Exploration is Harder than Prediction: Cryptographically Separating Reinforcement Learning from Supervised Learning.

Noah Golowich, Ankur Moitra, Dhruv Rohatgi

VenueA*FOCS
Year2024
ProceedingsFOCS

Browse the full FOCS paper archive.