Scalable Extraction of Training Data from Aligned, Production Language Models.
Milad Nasr, Javier Rando, Nicholas Carlini, Jonathan Hayase, Matthew Jagielski, A. Feder Cooper, Daphne Ippolito, Christopher A. Choquette-Choo, Florian Tramr, Katherine Lee
Browse the full ICLR paper archive.