Skip to content

Offline Reinforcement Learning via Policy Regularization and Ensemble Q-Functions.

Tao Wang, Shaorong Xie, Mingke Gao, Xue Chen, Zhenyu Zhang, Hang Yu

VenueBICTAI
Year2022
ProceedingsICTAI

Browse the full ICTAI paper archive.