2 papers
cs.RO2026
GesVLA: Gesture-Aware Vision-Language-Action Model Embedded Representations
Wenxuan Guo, Ziyuan Li, Meng Zhang +7
Vision-Language-Action (VLA) models have shown strong potential for general-purpose robot manipulation by unifying perception and action. However, existing VLA systems primarily re…
cs.CV2024
UniHands: Unifying Various Wild-Collected Keypoints for Personalized Hand Reconstruction
Menghe Zhang, Joonyeoup Kim, Yangwen Liang +2
Accurate hand motion capture and standardized 3D representation are essential for various hand-related tasks. Collecting keypoints-only data, while efficient and cost-effective, re…