works on

From the 2 of 69 linked papers with an AI index.

activity
20242026
collaborators
Showing cs.ROShow all

5 papers · 1 filter

cs.RO2026

See like a Robot: Robot-Centric Pointmaps for Vision-Language-Action Models

Byungkun Lee, Dongyoon Hwang, Dongjin Kim +3

The paper proposes robot-centric pointmaps, which encode 3D scene coordinates in the robot's frame as image pixels, enabling vision‑language‑action models to align visual inputs wi…

cs.RO2026

3D HAMSTER: Bridging Planning and Control in Hierarchical Vision Language Action Models through 3D Trajectory Guidance

Dongyoon Hwang, Byungkun Lee, Dongjin Kim +7

Hierarchical Vision-Language-Action (VLA) models decouple high-level planning from low-level control to improve generalization in robot manipulation. Recent work in this paradigm u…

cs.RO2026

Object-Centric Residual RL for Zero-Shot Sim-to-Real VLA Enhancement

Kinam Kim, Namiko Saito, Heecheol Kim +3

Vision-Language-Action (VLA) models can generalize across diverse manipulation tasks, but their imitation-learning-based policies remain brittle in precise physical interactions du…

cs.RO2026

PHUMA: Physically Reliable Humanoid Locomotion Dataset

Kyungmin Lee, Sibeen Kim, Youngdo Lee +6

Motion imitation is a promising approach for humanoid locomotion, enabling agents to acquire humanlike behaviors. Existing methods typically rely on high-quality motion capture dat…

cs.RO2026

ACG: Action Coherence Guidance for Flow-based Vision-Language-Action models

Minho Park, Kinam Kim, Junha Hyung +5

Diffusion and flow matching models have emerged as powerful robot policies, enabling Vision-Language-Action (VLA) models to generalize across diverse scenes and instructions. Yet,…