6.1k citations · 6.2k across the 22 of their papers we have counts for
1 paper · 1 filter
Sanchit Sinha, Oana Frunza, Kashif Rasul +2
The capabilities of Large Vision-Language Models (LVLMs) have reached state-of-the-art on many visual reasoning tasks, including chart reasoning, yet they still falter on out-of-di…