Skip to content

Beyond Correctness: Confidence-Aware Reward Modeling for Enhancing Large Language Model Reasoning.

Qianxi He, Qingyu Ren, Shanzhe Lei, Xuhong Wang, Yingchun Wang

VenueA*EMNLP
Year2025
ProceedingsEMNLP

Browse the full EMNLP paper archive.