Skip to content

Non-Linear Reinforcement Learning in Large Action Spaces: Structural Conditions and Sample-efficiency of Posterior Sampling.

Alekh Agarwal, Tong Zhang

VenueA*COLT
Year2022
ProceedingsCOLT

Browse the full COLT paper archive.