2 papers
cs.CV2025
Test-Time Consistency in Vision Language Models
Shih-Han Chou, Shivam Chandhok, James J. Little +1
Vision-Language Models (VLMs) have achieved impressive performance across a wide range of multimodal tasks, yet they often exhibit inconsistent behavior when faced with semanticall…
cs.CV2025
MM-R: On (In-)Consistency of Vision-Language Models (VLMs)
Shih-Han Chou, Shivam Chandhok, James J. Little +1
With the advent of LLMs and variants, a flurry of research has emerged, analyzing the performance of such models across an array of tasks. While most studies focus on evaluating th…