1 paper
Seongmin Jung, Seongho Choi, Gunwoo Jeon +2
3D Visual Grounding (3DVG) is a critical bridge from vision-language perception to robotics, requiring both language understanding and 3D scene reasoning. Traditional supervised mo…