most citedViTacFormer: Learning Cross-Modal Representation for Visuo-Tactile Dexterous Manipulation

1 citations · 1 across the 8 of their papers we have counts for

collaborators
Showing cs.ROShow all

12 papers · 1 filter

cs.RO2026

DIPOLE: Fusing Vision and Geometry for Robust Visuomotor Generalization

Yikai Tang, Haoran Geng, Jindou Jia +5

Imitation learning has emerged as a crucial approach for acquiring visuomotor skills from demonstrations, where designing effective observation encoders is essential for policy gen…

cs.RO20261 cited

ViTacFormer: Learning Cross-Modal Representation for Visuo-Tactile Dexterous Manipulation

Liang Heng, Haoran Geng, Kaifeng Zhang +2

Dexterous manipulation is a cornerstone capability for robotic systems aiming to interact with the physical world in a human-like manner. Although vision-based methods have advance…

cs.RO2026

Large Video Planner Enables Generalizable Robot Control

Boyuan Chen, Tianyuan Zhang, Haoran Geng +9

General-purpose robots require decision-making models that generalize across diverse tasks and environments. Recent works build robot foundation models by extending multimodal larg…

cs.RO2026

World Model for Robot Learning: A Comprehensive Survey

Bohan Hou, Gen Li, Jindou Jia +15

World models, which are predictive representations of how environments evolve under actions, have become a central component of robot learning. They support policy learning, planni…

cs.RO2026

Rodrigues Network for Learning Robot Actions

Jialiang Zhang, Haoran Geng, Yang You +4

Understanding and predicting articulated actions is important in robot learning. However, common architectures such as MLPs and Transformers lack inductive biases that reflect the…

cs.RO2026

D-REX: Differentiable Real-to-Sim-to-Real Engine for Learning Dexterous Grasping

Haozhe Lou, Mingtong Zhang, Haoran Geng +9

Simulation provides a cost-effective and flexible platform for data generation and policy learning to develop robotic systems. However, bridging the gap between simulation and real…