Skip to content

Learning Self-Correctable Policies and Value Functions from Demonstrations with Negative Sampling.

Yuping Luo, Huazhe Xu, Tengyu Ma

VenueA*ICLR
Year2020
ProceedingsICLR

Browse the full ICLR paper archive.