Skip to content

An Efficient and Precise Training Data Construction Framework for Process-supervised Reward Model in Mathematical Reasoning.

Wei Sun, Qianlong Du, Fuwei Cui, Jiajun Zhang

VenueA*ACL
Year2025
ProceedingsACL (1)

Browse the full ACL paper archive.