Skip to content

Efficient Arbitrary Precision Acceleration for Large Language Models on GPU Tensor Cores.

Shaobo Ma, Chao Fang, Haikuo Shao, Zhongfeng Wang

VenueBASPDAC
Year2025
ProceedingsASP-DAC

Browse the full ASPDAC paper archive.