Emerging Safety Attack and Defense in Federated Instruction Tuning of Large Language Models.
Rui Ye, Jingyi Chai, Xiangrui Liu, Yaodong Yang, Yanfeng Wang, Siheng Chen
Browse the full ICLR paper archive.
Rui Ye, Jingyi Chai, Xiangrui Liu, Yaodong Yang, Yanfeng Wang, Siheng Chen
Browse the full ICLR paper archive.