Skip to content

PELM: Power Efficient On-Device LLM Inference with Speculative Decoding and Dynamic Voltage Frequency Scaling.

Weisi Yang, Stephen Xia

Year2026
ProceedingsSenSys

Browse the full SENSYS paper archive.