1 paper · 1 filter
Kushin Mukherjee, Donghao Ren, Dominik Moritz +1
Multimodal vision-language models (VLMs) continue to achieve ever-improving scores on chart understanding benchmarks. Yet, we find that this progress does not fully capture the bre…