Skip to content

NADPEx: An on-policy temporally consistent exploration method for deep reinforcement learning.

Sirui Xie, Junning Huang, Lanxin Lei, Chunxiao Liu, Zheng Ma, Wei Zhang, Liang Lin

VenueA*ICLR
Year2019
ProceedingsICLR (Poster)

Browse the full ICLR paper archive.