Skip to content

throttLL'eM: Predictive GPU Throttling for Energy Efficient LLM Inference Serving.

Andreas Kosmas Kakolyris, Dimosthenis Masouros, Petros Vavaroutsos, Sotirios Xydis, Dimitrios Soudris

VenueA*HPCA
Year2025
ProceedingsHPCA

Browse the full HPCA paper archive.