1 paper
Jiasheng Li, Zhong Ji, Yan Zhang +1
Large Vision-Language Models (VLMs) suffer from prohibitive inference overhead due to long sequences of visual tokens. However, existing visual token reduction methods mainly impro…