Skip to content

Self-play with Execution Feedback: Improving Instruction-following Capabilities of Large Language Models.

Guanting Dong, Keming Lu, Chengpeng Li, Tingyu Xia, Bowen Yu, Chang Zhou, Jingren Zhou

VenueA*ICLR
Year2025
ProceedingsICLR

Browse the full ICLR paper archive.