Skip to content

Encouraging Good Processes Without the Need for Good Answers: Reinforcement Learning for LLM Agent Planning.

Zhiwei Li, Yong Hu, Wenqing Wang

VenueA*EMNLP
Year2025
ProceedingsEMNLP (Industry Track)

Browse the full EMNLP paper archive.