2 papers
cs.CV2026
EgoGenesis: Egocentric World-Action Modeling with Online Anchored Projective Memory and Action-3D RoPE
Zexuan Yan, Yuzhou Wu, Yue Ma +9
Egocentric video offers rich manipulation experience for embodied AI, yet collecting diverse egocentric data across scenes, objects, motions, and embodiments remains costly. We pre…
cs.RO2023
QwenGrasp: A Usage of Large Vision-Language Model for Target-Oriented Grasping
Xinyu Chen, Jian Yang, Zonghan He +3
Target-oriented grasping in unstructured scenes with language control is essential for intelligent robot arm grasping. The ability for the robot arm to understand the human languag…