1 paper · 1 filter
René Peinl, Vincent Tischler
This paper introduces a novel benchmark dataset designed to evaluate the capabilities of Vision Language Models (VLMs) on tasks that combine visual reasoning with subject-specific…