Skip to content

A Minimaximalist Approach to Reinforcement Learning from Human Feedback.

Gokul Swamy, Christoph Dann, Rahul Kidambi, Steven Wu, Alekh Agarwal

VenueA*ICML
Year2024
ProceedingsICML

Browse the full ICML paper archive.