Skip to content

Minimax-optimal reward-agnostic exploration in reinforcement learning.

Gen Li, Yuling Yan, Yuxin Chen, Jianqing Fan

VenueA*COLT
Year2024
ProceedingsCOLT

Browse the full COLT paper archive.