Efficient Online Reinforcement Learning Fine-Tuning Need Not Retain Offline Data.
Zhiyuan Zhou, Andy Peng, Qiyang Li, Sergey Levine, Aviral Kumar
Browse the full ICLR paper archive.
Zhiyuan Zhou, Andy Peng, Qiyang Li, Sergey Levine, Aviral Kumar
Browse the full ICLR paper archive.