Provably Efficient Exploration in Quantum Reinforcement Learning with Logarithmic Worst-Case Regret.
Han Zhong, Jiachen Hu, Yecheng Xue, Tongyang Li, Liwei Wang
Browse the full ICML paper archive.
Han Zhong, Jiachen Hu, Yecheng Xue, Tongyang Li, Liwei Wang
Browse the full ICML paper archive.