Enhancing Computation Efficiency in Large Language Models through Weight and Activation Quantization.
Janghwan Lee, Minsoo Kim, Seungcheol Baek, Seok Joong Hwang, Wonyong Sung, Jungwook Choi
Browse the full EMNLP paper archive.
Janghwan Lee, Minsoo Kim, Seungcheol Baek, Seok Joong Hwang, Wonyong Sung, Jungwook Choi
Browse the full EMNLP paper archive.