Skip to content

Alignment-Enhanced Decoding: Defending Jailbreaks via Token-Level Adaptive Refining of Probability Distributions.

Quan Liu, Zhenhong Zhou, Longzhu He, Yi Liu, Wei Zhang, Sen Su

VenueA*EMNLP
Year2024
ProceedingsEMNLP

Browse the full EMNLP paper archive.