Skip to content

Prefixing Attention Sinks can Mitigate Activation Outliers for Large Language Model Quantization.

Seungwoo Son, Wonpyo Park, Woohyun Han, Kyuyeun Kim, Jaeho Lee

VenueA*EMNLP
Year2024
ProceedingsEMNLP

Browse the full EMNLP paper archive.