1 paper
Eunsoo Lee, Jeongwoo Lee, Minki Hong +2
Large vision-language models (LVLMs) struggle to reliably detect visual primitives in charts and align them with semantic representations, which severely limits their performance o…