1 paper · 1 filter
Yuyi Li, Daoyuan Chen, Zhen Wang +2
Large Vision-Language Models (LVLMs) show promise for scientific applications, yet open-source models still struggle with Scientific Visual Question Answering (SVQA), namely answer…