What Kind of Visual Tokens Do We Need? Training-Free Visual Token Pruning for Multi-Modal Large Language Models from the Perspective of Graph.
Yutao Jiang, Qiong Wu, Wenhao Lin, Wei Yu, Yiyi Zhou
Browse the full AAAI paper archive.
Yutao Jiang, Qiong Wu, Wenhao Lin, Wei Yu, Yiyi Zhou
Browse the full AAAI paper archive.