Skip to content

On the Robustness of Reward Models for Language Model Alignment.

Jiwoo Hong, Noah Lee, Eunki Kim, Guijin Son, Woojin Chung, Aman Gupta, Shao Tang, James Thorne

VenueA*ICML
Year2025
ProceedingsICML

Browse the full ICML paper archive.