Skip to content

Explicit Preference Optimization: No Need for an Implicit Reward Model.

Xiangkun Hu, Lemin Kong, Tong He, David Wipf

VenueA*ICML
Year2025
ProceedingsICML

Browse the full ICML paper archive.