Skip to content

Enhancing Reinforcement Learning with Dense Rewards from Language Model Critic.

Meng Cao, Lei Shu, Lei Yu, Yun Zhu, Nevan Wichers, Yinxiao Liu, Lei Meng

VenueA*EMNLP
Year2024
ProceedingsEMNLP

Browse the full EMNLP paper archive.