1 paper
René Peinl, Vincent Tischler, Patrick Schröder +1
We present SITUATE, a novel dataset designed for training and evaluating Vision Language Models on counting tasks with spatial constraints. The dataset bridges the gap between simp…