Skip to content

LUT-LLM: Efficient Language Model Inference with Memory-based Computations on FPGAs.

Zifan He, Shengyu Ye, Rui Ma, Yang Wang, Jason Cong

Year2026
ProceedingsFCCM

Browse the full FCCM paper archive.