Flora: Sample-Efficient Preference-Based Rl Via Low-Rank Style Adaptation of Reward Functions.
Daniel Marta, Simon Holk, Miguel Vasco, Jens Lundell, Timon Homberger, Finn Busch, Olov Andersson, Danica Kragic, Iolanda Leite
Browse the full ICRA paper archive.