1 paper
Haotian Li, Yida Wang, Leyuan Wang +7
In recent years, multimodal large language models (MLLMs) have shown strong potential for embodied intelligence, yet their ability to maintain geometrically consistent spatial unde…