1 paper
Jingyu Lei, Gaoang Wang, Der-Horng Lee
Large Vision-Language Models (LVLMs) usually suffer from prohibitive computational and memory costs due to the quadratic growth of visual tokens with image resolution. Existing tok…