Skip to content

HookMoE: A learnable performance compensation strategy of Mixture-of-Experts for LLM inference acceleration.

Longkai Cheng, Along He, Mulin Li, Xueshuo Xie, Tao Li

VenueA*EMNLP
Year2025
ProceedingsEMNLP

Browse the full EMNLP paper archive.