3 papers
cs.RO2026
GraspLDP: Towards Generalizable Grasping Policy via Latent Diffusion
Enda Xiang, Haoxiang Ma, Xinzhu Ma +2
This paper focuses on enhancing the grasping precision and generalization of manipulation policies learned via imitation learning. Diffusion-based policy learning methods have rece…
cs.CV2025
Point2Primitive: CAD Reconstruction from Point Cloud by Direct Primitive Prediction
Xinzhu Ma, Cheng Wang, Chen Tang +5
Recovering CAD models from point clouds requires reconstructing their topology and sketch-based extrusion primitives. A dominant paradigm for representing sketches involves implici…
cs.CV2025
3DAxisPrompt: Promoting the 3D Grounding and Reasoning in GPT-4o
Dingning Liu, Cheng Wang, Peng Gao +4
Multimodal Large Language Models (MLLMs) exhibit impressive capabilities across a variety of tasks, especially when equipped with carefully designed visual prompts. However, existi…