Skip to content

Reinforcement Learning-powered Effectiveness and Efficiency Few-shot Jailbreaking Attack LLMs.

Xuehai Tang, Zhongjiang Yao, Jie Wen, Yangchen Dong, Jizhong Han, Songlin Hu

VenueCISPA
Year2024
ProceedingsISPA

Browse the full ISPA paper archive.