Approximation Error Back-Propagation for Q-Function in Scalable Reinforcement Learning with Tree Dependence Structure.
Yuzi Yan, Yu Dong, Kai Ma, Yuan Shen
Browse the full ICASSP paper archive.
Yuzi Yan, Yu Dong, Kai Ma, Yuan Shen
Browse the full ICASSP paper archive.