An Optimal Discriminator Weighted Imitation Perspective for Reinforcement Learning.
Haoran Xu, Shuozhe Li, Harshit Sikchi, Scott Niekum, Amy Zhang
Browse the full ICLR paper archive.
Haoran Xu, Shuozhe Li, Harshit Sikchi, Scott Niekum, Amy Zhang
Browse the full ICLR paper archive.