1 paper
He Liang, Chenyang Ma, Yiming Zhang +4
Existing 3D scene-grounded Large Language Models (3D-LLMs) focus on answering questions grounded in simplified single-room 3D scenes, lacking the ability to reason over real-world…