Skip to content

ROSE Doesn't Do That: Boosting the Safety of Instruction-Tuned Large Language Models with Reverse Prompt Contrastive Decoding.

Qihuang Zhong, Liang Ding, Juhua Liu, Bo Du, Dacheng Tao

VenueA*ACL
Year2024
ProceedingsACL (Findings)

Browse the full ACL paper archive.