Skip to content

WildReward: Learning Reward Models from In-the-Wild Human Interactions.

Hao Peng, Yunjia Qi, Xiaozhi Wang, Zijun Yao, Lei Hou, Juanzi Li

VenueA*ACL
Year2026
ProceedingsACL (1)

Browse the full ACL paper archive.