1 paper · 1 filter
Kai Sun, Yushi Bai, Ji Qi +2
To advance the evaluation of multimodal math reasoning in large multimodal models (LMMs), this paper introduces a novel benchmark, MM-MATH. MM-MATH consists of 5,929 open-ended mid…