REINFORCE Adversarial Attacks on Large Language Models: An Adaptive, Distributional, and Semantic Objective.
Simon Geisler, Tom Wollschlger, M. H. I. Abdalla, Vincent Cohen-Addad, Johannes Gasteiger, Stephan Gnnemann
Browse the full ICML paper archive.