Skip to content

RLCD: Reinforcement Learning from Contrastive Distillation for LM Alignment.

Kevin Yang, Dan Klein, Asli Celikyilmaz, Nanyun Peng, Yuandong Tian

VenueA*ICLR
Year2024
ProceedingsICLR

Browse the full ICLR paper archive.