Skip to content

Train Flat, Then Compress: Sharpness-Aware Minimization Learns More Compressible Models.

Clara Na, Sanket Vaibhav Mehta, Emma Strubell

VenueA*EMNLP
Year2022
ProceedingsEMNLP (Findings)

Browse the full EMNLP paper archive.