Skip to content

TreeRL: LLM Reinforcement Learning with On-Policy Tree Search.

Zhenyu Hou, Ziniu Hu, Yujiang Li, Rui Lu, Jie Tang, Yuxiao Dong

VenueA*ACL
Year2025
ProceedingsACL (1)

Browse the full ACL paper archive.