When Prompt Optimization Becomes Jailbreaking: Adaptive Red-Teaming of Large Language Models.
Zafir Shamsi, Nikhil Chekuru, Zachary Guzman, Shivank Garg
Browse the full EACL paper archive.
Zafir Shamsi, Nikhil Chekuru, Zachary Guzman, Shivank Garg
Browse the full EACL paper archive.