1 paper · 1 filter
Yingbing Huang, Tharun Adithya Srikrishnan, Steven K. Reinhardt +1
Vision-Language Models (VLMs) have emerged as a critical and fast-growing extension of Large Language Models (LLMs) that enable multimodal reasoning through both text and image inp…