Skip to content

Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction.

Zhenmei Shi, Yifei Ming, Xuan-Phi Nguyen, Yingyu Liang, Shafiq Joty

VenueA*ACL
Year2026
ProceedingsACL (Findings)

Browse the full ACL paper archive.