Skip to content

Regularizing a Model-based Policy Stationary Distribution to Stabilize Offline Reinforcement Learning.

Shentao Yang, Yihao Feng, Shujian Zhang, Mingyuan Zhou

VenueA*ICML
Year2022
ProceedingsICML

Browse the full ICML paper archive.