Skip to content

Efficiently Learning To Reason or Not to Reason: Root-token Policy Optimization for Adaptive Thinking.

Taehyeon Kim, Hyunsoo Lee, Youngsoo Jang, Moontae Lee

VenueA*ACL
Year2026
ProceedingsACL (1)

Browse the full ACL paper archive.