Skip to content

DOPT: D-Learning with Off-Policy Target toward Sample Efficiency and Fast Convergence Control.

Zhaolong Shen, Quan Quan

VenueA*ICRA
Year2025
ProceedingsICRA

Browse the full ICRA paper archive.