1 paper · 1 filter
Yitong Zhou, Mingyue Cheng, Qingyang Mao +7
With the widespread application of multimodal large language models in scientific intelligence, there is an urgent need for more challenging evaluation benchmarks to assess their a…