1 paper · 1 filter
Mohit Vaishnav, Tanel Tammet
Vision--language models (VLMs) often fail on abstract visual reasoning benchmarks such as Bongard problems, raising the question of whether the main bottleneck lies in reasoning or…