From the 1 of 16 linked papers with an AI index.
1 citations · 1 across the 9 of their papers we have counts for
4 papers · 1 filter
See like a Robot: Robot-Centric Pointmaps for Vision-Language-Action Models
Byungkun Lee, Dongyoon Hwang, Dongjin Kim +3
The paper proposes robot-centric pointmaps, which encode 3D scene coordinates in the robot's frame as image pixels, enabling vision‑language‑action models to align visual inputs wi…
3D HAMSTER: Bridging Planning and Control in Hierarchical Vision Language Action Models through 3D Trajectory Guidance
Dongyoon Hwang, Byungkun Lee, Dongjin Kim +7
Hierarchical Vision-Language-Action (VLA) models decouple high-level planning from low-level control to improve generalization in robot manipulation. Recent work in this paradigm u…
PHUMA: Physically Reliable Humanoid Locomotion Dataset
Kyungmin Lee, Sibeen Kim, Youngdo Lee +6
Motion imitation is a promising approach for humanoid locomotion, enabling agents to acquire humanlike behaviors. Existing methods typically rely on high-quality motion capture dat…
ACG: Action Coherence Guidance for Flow-based Vision-Language-Action models
Minho Park, Kinam Kim, Junha Hyung +5
Diffusion and flow matching models have emerged as powerful robot policies, enabling Vision-Language-Action (VLA) models to generalize across diverse scenes and instructions. Yet,…