ConcaveQ: Non-monotonic Value Function Factorization via Concave Representations in Deep Multi-Agent Reinforcement Learning.
Huiqun Li, Hanhan Zhou, Yifei Zou, Dongxiao Yu, Tian Lan
Browse the full AAAI paper archive.
Huiqun Li, Hanhan Zhou, Yifei Zou, Dongxiao Yu, Tian Lan
Browse the full AAAI paper archive.