Skip to content

Fine-tuning Reinforcement Learning Models is Secretly a Forgetting Mitigation Problem.

Maciej Wolczyk, Bartlomiej Cupial, Mateusz Ostaszewski, Michal Bortkiewicz, Michal Zajac, Razvan Pascanu, Lukasz Kucinski, Piotr Milos

VenueA*ICML
Year2024
ProceedingsICML

Browse the full ICML paper archive.