Fast and Data Efficient Reinforcement Learning from Pixels via Non-parametric Value Approximation.
Alexander Long, Alan Blair, Herke van Hoof
Browse the full AAAI paper archive.
Alexander Long, Alan Blair, Herke van Hoof
Browse the full AAAI paper archive.