1 paper
Yang Yang, Jiawei Chen, Tairan Chen +1
Although Multimodal Large Language Models (MLLMs) have made substantial progress, their spatial reasoning may still produce intermediate judgments inconsistent with the input image…