1 paper
Jason Qiu, Zachary Meurer, Xavier Thomas +1
This work investigates the fundamental fragility of state-of-the-art Vision-Language Models (VLMs) under basic geometric transformations. While modern VLMs excel at semantic tasks…