5 papers
Moving Out: Physically-grounded Human-AI Collaboration
Xuhui Kang, Sung-Wook Lee, Haolin Liu +2
The ability to adapt to physical actions and constraints in an environment is crucial for embodied agents (e.g., robots) to effectively collaborate with humans. Such physically gro…
TBD-VLA: Temporal Block Diffusion Vision Language Action Model
Sung-Wook Lee, Xuhui Kang, Yen-Ling Kuo
Discrete Vision-Language-Action (VLA) models typically formulate action generation as next-token prediction over discretized action spaces, conditioning each token autoregressively…
Learning Force-Regulated Manipulation with a Low-Cost Tactile-Force-Controlled Gripper
Xuhui Kang, Tongxuan Tian, Sung-Wook Lee +3
Successfully manipulating many everyday objects, such as potato chips, requires precise force regulation. Failure to modulate force can lead to task failure or irreversible damage…
CLASS: Contrastive Learning via Action Sequence Supervision for Robot Manipulation
Sung-Wook Lee, Xuhui Kang, Brandon Yang +1
Recent advances in Behavior Cloning (BC) have led to strong performance in robotic manipulation, driven by expressive models, sequence modeling of actions, and large-scale demonstr…
Diff-DAgger: Uncertainty Estimation with Diffusion Policy for Robotic Manipulation
Sung-Wook Lee, Xuhui Kang, Yen-Ling Kuo
Recently, diffusion policy has shown impressive results in handling multi-modal tasks in robotic manipulation. However, it has fundamental limitations in out-of-distribution failur…