Skip to content

Boost Transformer-based Language Models with GPU-Friendly Sparsity and Quantization.

Chong Yu, Tao Chen, Zhongxue Gan

VenueA*ACL
Year2023
ProceedingsACL (Findings)

Browse the full ACL paper archive.