1 paper
Haida Feng, Hao Wei, Haolin Wang +3
Recent Multimodal Large Language Models (MLLMs) struggle to bridge the representational gap between 2D semantic understanding and 3D spatial geometry. Existing 3D-aware models eith…