Skip to content

Learning Infinite-horizon Average-reward MDPs with Linear Function Approximation.

Chen-Yu Wei, Mehdi Jafarnia-Jahromi, Haipeng Luo, Rahul Jain

Year2021
ProceedingsAISTATS

Browse the full AISTATS paper archive.