1 paper
Shan Zhang, Aotian Chen, Yanpeng Sun +6
Current multimodal large language models (MLLMs) often underperform on mathematical problem-solving tasks that require fine-grained visual understanding. The limitation is largely…