Skip to content

Sample-efficient Learning of Infinite-horizon Average-reward MDPs with General Function Approximation.

Jianliang He, Han Zhong, Zhuoran Yang

VenueA*ICLR
Year2024
ProceedingsICLR

Browse the full ICLR paper archive.