Learning from Demonstration: Provably Efficient Adversarial Policy Imitation with Linear Function Approximation.
Zhihan Liu, Yufeng Zhang, Zuyue Fu, Zhuoran Yang, Zhaoran Wang
Browse the full ICML paper archive.
Zhihan Liu, Yufeng Zhang, Zuyue Fu, Zhuoran Yang, Zhaoran Wang
Browse the full ICML paper archive.