1 citations · 1 across the 6 of their papers we have counts for
8 papers
Assistron: Bayesian Shared Autonomy with Off-the-shelf Vision-Language-Action Models
Pinhao Song, Ze Fu, Yutong Hu +1
We propose Assistron, a shared autonomy model that leverages Vision-Language-Action (VLA) models to assist the user in daily activities. Our approach is grounded in two core princi…
TASC: Task-Aware Shared Control for Relational Telemanipulation
Ze Fu, Pinhao Song, Yutong Hu +1
We present TASC, a Task-Aware Shared Control framework for relational telemanipulation that infers task-level user intent and provides assistance from motion-only input. To support…
FF-JEPA: Long-Horizon Planning in World Models with Latent Planners
Sergi Masip, Jonathan Swinnen, Yutong Hu +2
Joint Embedding Predictive Architectures (JEPAs) have shown promising world modeling capabilities, enabling planning in latent space by optimizing action trajectories using methods…
AR-VLA: True Autoregressive Action Expert for Vision-Language-Action Models
Yutong Hu, Jan-Nico Zaech, Nikolay Nikolov +6
We propose a standalone autoregressive (AR) Action Expert that generates actions as a continuous causal sequence while conditioning on refreshable vision-language prefixes. In cont…
Equivariant Volumetric Grasping
Pinhao Song, Yutong Hu, Pengteng Li +1
We propose a new volumetric grasp model that is equivariant to rotations around the vertical axis, leading to a significant improvement in sampling efficiency. Our model employs a…
ELVIS: Ensemble-Calibrated Latent Imagination for Long-Horizon Visual MPC
Yurui Du, Pinhao Song, Yutong Hu +1
A central challenge of visual control with model-based reinforcement learning (RL) is reliable long-horizon planning: long rollouts with learned latent dynamics exhibit branching f…