Skip to content

Human-in-the-loop: Provably Efficient Preference-based Reinforcement Learning with General Function Approximation.

Xiaoyu Chen, Han Zhong, Zhuoran Yang, Zhaoran Wang, Liwei Wang

VenueA*ICML
Year2022
ProceedingsICML

Browse the full ICML paper archive.