2 papers
cs.AI2026
VisDoT : Enhancing Visual Reasoning through Human-Like Interpretation Grounding and Decomposition of Thought
Eunsoo Lee, Jeongwoo Lee, Minki Hong +2
Large vision-language models (LVLMs) struggle to reliably detect visual primitives in charts and align them with semantic representations, which severely limits their performance o…
cs.AI2026
Hospitality-VQA: Decision-Oriented Informativeness Evaluation for Vision-Language Models
Jeongwoo Lee, Baek Duhyeong, Eungyeol Han +5
Recent advances in Vision-Language Models (VLMs) have demonstrated impressive multimodal understanding in general domains. However, their applicability to decision-oriented domains…