SAEIR: Sequentially Accumulated Entropy Intrinsic Reward for Cooperative Multi-Agent Reinforcement Learning with Sparse Reward.
Xin He, Hongwei Ge, Yaqing Hou, Jincheng Yu
Browse the full IJCAI paper archive.
Xin He, Hongwei Ge, Yaqing Hou, Jincheng Yu
Browse the full IJCAI paper archive.