1 paper
Kanghee Lee, Jungi Hong, Sion Lee +4
Recent progress in Multimodal Large Language Models (MLLMs) has enabled 3D scene understanding and spatial reasoning directly from multi-view images, without requiring explicit 3D…