3 papers
cs.CV2026
SceneGraphGrounder: Zero-Shot 3D Visual Grounding via Structured Scene Graph Matching
Xuefei Sun, Xujia Zhang, Brendan Crowe +2
Zero-shot 3D visual grounding requires localizing objects in unstructured environments from free-form natural language. Recent vision-language model (VLM) approaches achieve promis…
cs.CV2025
Octree Diffusion for Semantic Scene Generation and Completion
Xujia Zhang, Brendan Crowe, Christoffer Heckman
The completion, extension, and generation of 3D semantic scenes are an interrelated set of capabilities that are useful for robotic navigation and exploration. Existing approaches…
cs.RO2023
Toward Optimal Tabletop Rearrangement with Multiple Manipulation Primitives
Baichuan Huang, Xujia Zhang, Jingjin Yu
In practice, many types of manipulation actions (e.g., pick-n-place and push) are needed to accomplish real-world manipulation tasks. Yet, limited research exists that explores the…