Vista-LLM: Decoupled Query-Guided Visual Token Pruning for Efficient Long-Video Large Language Models.
Zhenyu Li, Zuchao Li, Ping Wang, Lefei Zhang, Haojun Ai
Browse the full ACL paper archive.
Zhenyu Li, Zuchao Li, Ping Wang, Lefei Zhang, Haojun Ai
Browse the full ACL paper archive.