1 paper
Mingze Yin, Xiaohan Wang, Dian Li +8
Performing deliberate mathematical reasoning in visual contexts is a hallmark of advanced Multimodal Large Language Models (MLLMs) and requires a sophisticated synthesis of percept…