1 paper
Nikita Drozdov, Andrey Lemeshko, Nikita Gavrilov +3
3D visual grounding (3DVG) aims to localize objects in a 3D scene based on natural language queries. In this work, we explore zero-shot 3DVG from multi-view images alone, without r…