1 paper · 1 filter
Eunsoo Lee, Jeongwoo Lee, Minki Hong +2
Large vision-language models (LVLMs) struggle to reliably detect visual primitives in charts and align them with semantic representations, which severely limits their performance o…