The Perils of Optimizing Learned Reward Functions: Low Training Error Does Not Guarantee Low Regret.
Lukas Fluri, Leon Lang, Alessandro Abate, Patrick Forr, David Krueger, Joar Max Viktor Skalse
Browse the full ICML paper archive.
Lukas Fluri, Leon Lang, Alessandro Abate, Patrick Forr, David Krueger, Joar Max Viktor Skalse
Browse the full ICML paper archive.