Booster: Tackling Harmful Fine-tuning for Large Language Models via Attenuating Harmful Perturbation.
Tiansheng Huang, Sihao Hu, Fatih Ilhan, Selim Furkan Tekin, Ling Liu
Browse the full ICLR paper archive.
Tiansheng Huang, Sihao Hu, Fatih Ilhan, Selim Furkan Tekin, Ling Liu
Browse the full ICLR paper archive.