2 papers
cs.RO2026
Dexterity-BEV: Aligning 3D World and Actions for Generalizable Robot Policies Learning
Huayi Zhou, Wei Gao, Dekun Lu +12
End-to-end manipulation policies, combined with web-scale pretrained Vision-Language Models (VLMs), show the promise for generalizable and dexterous robotic manipulation. However,…
cs.CV2024
DC3DO: Diffusion Classifier for 3D Objects
Nursena Koprucu, Meher Shashwat Nigam, Shicheng Xu +6
Inspired by Geoffrey Hinton emphasis on generative modeling, To recognize shapes, first learn to generate them, we explore the use of 3D diffusion models for object classification.…