Skip to content

BESA: Pruning Large Language Models with Blockwise Parameter-Efficient Sparsity Allocation.

Peng Xu, Wenqi Shao, Mengzhao Chen, Shitao Tang, Kaipeng Zhang, Peng Gao, Fengwei An, Yu Qiao, Ping Luo

VenueA*ICLR
Year2024
ProceedingsICLR

Browse the full ICLR paper archive.