1 paper
Yiwen Wang, Yuyang Deng, Yihao Long +1
3D visual grounding aims to localize the target object in a 3D scene from a natural language query, requiring both fine-grained semantic understanding and viewpoint-dependent spati…