Skip to content

M³HF: Multi-agent Reinforcement Learning from Multi-phase Human Feedback of Mixed Quality.

Ziyan Wang, Zhicheng Zhang, Fei Fang, Yali Du

VenueA*ICML
Year2025
ProceedingsICML

Browse the full ICML paper archive.