Skip to content

CSRP: Chain-of-Thought Reasoning for Chinese Text Correction via Reinforcement Learning with Efficiency-Aware Rewards.

Wei Tian, Yuhao Zhou, Man Lan

VenueA*ACL
Year2026
ProceedingsACL (1)

Browse the full ACL paper archive.