Skip to content

Learning Infinite-Horizon Average-Reward Linear Mixture MDPs of Bounded Span.

Woojin Chae, Kihyuk Hong, Yufan Zhang, Ambuj Tewari, Dabeen Lee

Year2025
ProceedingsAISTATS

Browse the full AISTATS paper archive.