Skip to content

Minimax Optimal Regret Bound for Reinforcement Learning with Trajectory Feedback.

Zihan Zhang, Yuxin Chen, Jason D. Lee, Simon Shaolei Du, Ruosong Wang

VenueA*ICML
Year2025
ProceedingsICML

Browse the full ICML paper archive.