Skip to content

Q-Prop: Sample-Efficient Policy Gradient with An Off-Policy Critic.

Shixiang Gu, Timothy P. Lillicrap, Zoubin Ghahramani, Richard E. Turner, Sergey Levine

VenueA*ICLR
Year2017
ProceedingsICLR

Browse the full ICLR paper archive.