Easy Samples Are All You Need: Self-Evolving LLMs via Data-Efficient Reinforcement Learning.
Zhiyin Yu, Bo Zhang, Qibin Hou, Zhonghai Wu, Xiao Luo, Lei Bai
Browse the full ACL paper archive.
Zhiyin Yu, Bo Zhang, Qibin Hou, Zhonghai Wu, Xiao Luo, Lei Bai
Browse the full ACL paper archive.