MMD-MIX: Value Function Factorisation with Maximum Mean Discrepancy for Cooperative Multi-Agent Reinforcement Learning.
Zhiwei Xu, Dapeng Li, Yunpeng Bai, Guoliang Fan
Browse the full IJCNN paper archive.
Zhiwei Xu, Dapeng Li, Yunpeng Bai, Guoliang Fan
Browse the full IJCNN paper archive.