1 paper · 1 filter
Daixian Liu, Jiayi Kuang, Yinghui Li +8
Multimodal Large Language Models (MLLMs) have achieved remarkable progress in visual recognition and semantic understanding, yet precise compositional spatial reasoning under geome…