Skip to content
cs-conference-ranking
.org
By subfield
By rank
Methodology
⌕
Search 971 venues
Home
/
ICML
/
Paper
Iterative Data Smoothing: Mitigating Reward Overfitting and Overoptimization in RLHF.
Banghua Zhu
,
Michael I. Jordan
,
Jiantao Jiao
Venue
A*
ICML
Year
2024
Proceedings
ICML
DBLP record
conf/icml/ZhuJJ24 ↗
Browse the full
ICML paper archive
.