2 papers
cs.CV2026
SceneGraphGrounder: Zero-Shot 3D Visual Grounding via Structured Scene Graph Matching
Xuefei Sun, Xujia Zhang, Brendan Crowe +2
Zero-shot 3D visual grounding requires localizing objects in unstructured environments from free-form natural language. Recent vision-language model (VLM) approaches achieve promis…
cs.CV2026
Octree Diffusion for Semantic Scene Generation and Completion
Xujia Zhang, Brendan Crowe, Christoffer Heckman
The completion, extension, and generation of 3D semantic scenes are an interrelated set of capabilities that are useful for robotic navigation and exploration. Existing approaches…