Extrapolating Beyond Suboptimal Demonstrations via Inverse Reinforcement Learning from Observations.
Daniel S. Brown, Wonjoon Goo, Prabhat Nagarajan, Scott Niekum
Browse the full ICML paper archive.
Daniel S. Brown, Wonjoon Goo, Prabhat Nagarajan, Scott Niekum
Browse the full ICML paper archive.