7 papers
GeoAware-VLA: Implicit Geometry Aware Vision-Language-Action Model
Ali Abouzeid, Malak Mansour, Qinbo Sun +2
Vision-Language-Action (VLA) models often fail to generalize to unseen camera viewpoints, a limitation stemming from their difficulty in inferring robust 3D geometry from 2D images…
UNCLE-Grasp: Uncertainty-Aware Grasping of Leaf-Occluded Strawberries
Malak Mansour, Ali Abouzeid, Zezhou Sun +3
Robotic strawberry harvesting remains challenging under partial occlusion, where leaf interference introduces significant geometric uncertainty and renders grasp decisions based on…
3D-CovDiffusion: 3D-Aware Diffusion Policy for Coverage Path Planning
Chenyuan Chen, Haoran Ding, Ran Ding +6
Diffusion models have shown strong potential for robot skill learning, yet their role in coverage path planning remains underexplored. In industrial surface processing (painting, p…
Imagination at Inference: Synthesizing In-Hand Views for Robust Visuomotor Policy Inference
Haoran Ding, Anqing Duan, Zezhou Sun +2
Visual observations from different viewpoints can significantly influence the performance of visuomotor policies in robotic manipulation. Among these, egocentric (in-hand) views of…
A Hybrid Hinge-Beam Continuum Robot with Passive Safety Capping for Real-Time Fatigue Awareness
Tongshun Chen, Zezhou Sun, Yanhan Sun +3
Cable-driven continuum robots offer high flexibility and lightweight design, making them well-suited for tasks in constrained and unstructured environments. However, prolonged use…
Towards Safe Imitation Learning via Potential Field-Guided Flow Matching
Haoran Ding, Anqing Duan, Zezhou Sun +4
Deep generative models, particularly diffusion and flow matching models, have recently shown remarkable potential in learning complex policies through imitation learning. However,…