Parallel Q-Learning: Scaling Off-policy Reinforcement Learning under Massively Parallel Simulation.
Zechu Li, Tao Chen, Zhang-Wei Hong, Anurag Ajay, Pulkit Agrawal
Browse the full ICML paper archive.
Zechu Li, Tao Chen, Zhang-Wei Hong, Anurag Ajay, Pulkit Agrawal
Browse the full ICML paper archive.