Skip to content

A reinforcement learning framework based on regret minimization for approximating best response in fictitious self-play.

Yanran Xu, Kangxin He, Shu Hu, Hui Li

VenueCHPCC
Year2022
ProceedingsHPCC/DSS/SmartCity/DependSys

Browse the full HPCC paper archive.