Skip to content
cs-conference-ranking
.org
By subfield
By rank
Methodology
⌕
Search 971 venues
Home
/
ICLR
/
Paper
SqueezeAttention: 2D Management of KV-Cache in LLM Inference via Layer-wise Optimal Budget.
Zihao Wang
,
Bin Cui
,
Shaoduo Gan
Venue
A*
ICLR
Year
2025
Proceedings
ICLR
DBLP record
conf/iclr/WangCG25 ↗
Browse the full
ICLR paper archive
.