3 papers
cs.CV2026
UnfoldArt: Zero-Shot Recovery of Full Articulated 3D Objects from Text or Image
Mohamed el Amine Boudjoghra, Ivan Laptev, Angela Dai
Articulated 3D objects are essential for interactive environments in embodied AI, robotics, and virtual reality, but reconstructing their structure and motion from sparse observati…
cs.RO2025
GLaD: Geometric Latent Distillation for Vision-Language-Action Models
Minghao Guo, Meng Cao, Jiachen Tao +5
Most existing Vision-Language-Action (VLA) models rely primarily on RGB information, while ignoring geometric cues crucial for spatial reasoning and manipulation. In this work, we…
cs.CV2025
ScanEdit: Hierarchically-Guided Functional 3D Scan Editing
Mohamed el amine Boudjoghra, Ivan Laptev, Angela Dai
With the fast pace of 3D capture technology and resulting abundance of 3D data, effective 3D scene editing becomes essential for a variety of graphics applications. In this work we…