1 paper
Jiayu Tang, Yuchen Zhou, Chao Gou
Unlocking the spatial intelligence of multimodal large language model (MLLMs) is crucial for understanding and interacting with the 3D world. Prevailing approaches typically inject…