34 citations · 39 across the 3 of their papers we have counts for
1 paper · 2 filters
Yida Xue, Zhen Bi, Jinnan Yang +5
Recent advances in Multimodal Large Language Models (MLLMs) have significantly enhanced their capabilities; however, their spatial perception abilities remain a notable limitation.…