4 papers
Pelican-VLA 0.5: Attending Before Acting Benefits Generalization
Zeyuan Ding, Wenhai Liu, Yang Xu +6
In this report, we present Pelican-VLA 0.5, a unified VLA model that integrates vision-language understanding, future-frame generation, and action prediction within a single archit…
Pelican-Unify 1.0: A Unified Embodied Intelligence Model for Understanding, Reasoning, Imagination and Action
Yi Zhang, Yinda Chen, Che Liu +26
We present Pelican-Unify 1.0, the first embodied foundation model trained according to the principle of unification. Pelican-Unify 1.0 uses a single VLM as a unified understanding…
ExoGS: A 4D Real-to-Sim-to-Real Framework for Scalable Manipulation Data Collection
Yiming Wang, Ruogu Zhang, Minyang Li +7
Real-to-Sim-to-Real technique is gaining increasing interest for robotic manipulation, as it can generate scalable data in simulation while having narrower sim-to-real gap. However…
ForceMimic: Force-Centric Imitation Learning with Force-Motion Capture System for Contact-Rich Manipulation
Wenhai Liu, Junbo Wang, Yiming Wang +2
In most contact-rich manipulation tasks, humans apply time-varying forces to the target object, compensating for inaccuracies in the vision-guided hand trajectory. However, current…