1 paper
Sifan Wu, Huan Zhang, Yizhan Li +3
The emergence of Multimodal Large Language Models (MLLMs) that integrate vision and language modalities has unlocked new potentials for scientific reasoning, outperforming prior be…