Skip to content

Dynamic Rewarding with Prompt Optimization Enables Tuning-free Self-Alignment of Language Models.

Somanshu Singla, Zhen Wang, Tianyang Liu, Abdullah Ashfaq, Zhiting Hu, Eric P. Xing

VenueA*EMNLP
Year2024
ProceedingsEMNLP

Browse the full EMNLP paper archive.