Skip to content

The Perils of Optimizing Learned Reward Functions: Low Training Error Does Not Guarantee Low Regret.

Lukas Fluri, Leon Lang, Alessandro Abate, Patrick Forr, David Krueger, Joar Max Viktor Skalse

VenueA*ICML
Year2025
ProceedingsICML

Browse the full ICML paper archive.