Skip to content

Investigating Acceleration of LLaMA Inference by Enabling Intermediate Layer Decoding via Instruction Tuning with 'LITE'.

Neeraj Varshney, Agneet Chatterjee, Mihir Parmar, Chitta Baral

VenueANAACL
Year2024
ProceedingsNAACL-HLT (Findings)

Browse the full NAACL paper archive.