Skip to content

Contrastive Preference Learning: Learning from Human Feedback without Reinforcement Learning.

Joey Hejna, Rafael Rafailov, Harshit Sikchi, Chelsea Finn, Scott Niekum, W. Bradley Knox, Dorsa Sadigh

VenueA*ICLR
Year2024
ProceedingsICLR

Browse the full ICLR paper archive.