1 paper · 1 filter
Sifan Wu, Huan Zhang, Yizhan Li +3
The emergence of Multimodal Large Language Models (MLLMs) that integrate vision and language modalities has unlocked new potentials for scientific reasoning, outperforming prior be…