1 paper · 1 filter
Aoming Liu, Reuben Tan, Boqing Gong +1
Token reduction accelerates Multimodal Large Language Models (MLLMs) by reducing excessive tokens, but overlooks structural redundancy differences, where critical and redundant mod…