4 papers
BridgeACT: Bridging Human Demonstrations to Robot Actions via Unified Tool-Target Affordances
Yifan Han, Jianxiang Liu, Haoyu Zhang +3
Learning robot manipulation from human videos is appealing due to the scale and diversity of human demonstrations, but transferring such demonstrations to executable robot behavior…
Facial Expression Generation Aligned with Human Preference for Natural Dyadic Interaction
Xu Chen, Rui Gao, Xinjie Zhang +5
Achieving natural dyadic interaction requires generating facial expressions that are emotionally appropriate and socially aligned with human preference. Human feedback offers a com…
Boosting Action-Information via a Variational Bottleneck on Unlabelled Robot Videos
Haoyu Zhang, Long Cheng
Learning from demonstrations (LfD) typically relies on large amounts of action-labeled expert trajectories, which fundamentally constrains the scale of available training data. A p…
Depth-PC: A Visual Servo Framework Integrated with Cross-Modality Fusion for Sim2Real Transfer
Haoyu Zhang, Yang Liu, Yimu Jiang +2
Visual servoing techniques guide robotic motion using visual information to accomplish manipulation tasks, requiring high precision and robustness against noise. Traditional method…