Skip to content

Degeneration-free Policy Optimization: RL Fine-Tuning for Language Models without Degeneration.

Youngsoo Jang, Geon-Hyeong Kim, Byoungjip Kim, Yu Jin Kim, Honglak Lee, Moontae Lee

VenueA*ICML
Year2024
ProceedingsICML

Browse the full ICML paper archive.