Skip to content

Near-optimal Policy Optimization Algorithms for Learning Adversarial Linear Mixture MDPs.

Jiafan He, Dongruo Zhou, Quanquan Gu

Year2022
ProceedingsAISTATS

Browse the full AISTATS paper archive.