Q Cache: Visual Attention Is Valuable in Less than Half of Decode Layers for Multimodal Large Language Model.
Jiedong Zhuang, Lu Lu, Ming Dai, Rui Hu, Jian Chen, Qiang Liu, Haoji Hu
Browse the full AAAI paper archive.
Jiedong Zhuang, Lu Lu, Ming Dai, Rui Hu, Jian Chen, Qiang Liu, Haoji Hu
Browse the full AAAI paper archive.