Skip to content

NIKA: Optimal KV Cache Transfer for Minimizing the Latency of Disaggregated LLM Inference.

Chih-Tai Tsai, Zhan-Wei Wu, Yi-Syuan Ke, Sao-Hsuan Lin, Jerry Chou

VenueCCLOSER
Year2026
ProceedingsCLOSER

Browse the full CLOSER paper archive.