1 paper · 1 filter
Surendra Pathak, Bo Han
Although Large Vision Language Models (LVLMs) have demonstrated impressive multimodal reasoning capabilities, their scalability and deployment are constrained by massive computatio…