Sample-Efficient Preference-based Reinforcement Learning with Dynamics Aware Rewards.
Katherine Metcalf, Miguel Sarabia, Natalie Mackraz, Barry-John Theobald
Browse the full CoRL paper archive.
Katherine Metcalf, Miguel Sarabia, Natalie Mackraz, Barry-John Theobald
Browse the full CoRL paper archive.