Skip to content

Rewards-in-Context: Multi-objective Alignment of Foundation Models with Dynamic Preference Adjustment.

Rui Yang, Xiaoman Pan, Feng Luo, Shuang Qiu, Han Zhong, Dong Yu, Jianshu Chen

VenueA*ICML
Year2024
ProceedingsICML

Browse the full ICML paper archive.