Response Attack: Exploiting Contextual Priming to Jailbreak Large Language Models.
Ziqi Miao, Lijun Li, Yuan Xiong, Zhenhua Liu, Pengyu Zhu, Jing Shao
Browse the full AAAI paper archive.
Ziqi Miao, Lijun Li, Yuan Xiong, Zhenhua Liu, Pengyu Zhu, Jing Shao
Browse the full AAAI paper archive.