1 paper · 1 filter
Chaofang Ma, Lin Jiang, Carol Jingyi Li +4
Vision-Language Models (VLMs) have exhibited impressive performance across diverse visual scenarios. However, this success comes at the cost of explosive growth in visual tokens, w…