Skip to content

Parallel Q-Learning: Scaling Off-policy Reinforcement Learning under Massively Parallel Simulation.

Zechu Li, Tao Chen, Zhang-Wei Hong, Anurag Ajay, Pulkit Agrawal

VenueA*ICML
Year2023
ProceedingsICML

Browse the full ICML paper archive.