5 citations · 5 across the 3 of their papers we have counts for
3 papers
cs.RO2026
ValueFormer: A Causal Transformer Value Function with Stage-Aware Labels for Semi-Autonomous Vision-Language-Action Policies
Inkyu Sa, Konstantin Stulov, Rajat Bhageria
Vision-Language-Action (VLA) policies trained by behavior cloning fail silently: from the action stream alone, a collapsing rollout looks much like one making clean progress, becau…
cs.RO2026
Vision Language Action (VLA) Models for Unmanned Aerial Robotics and Bimanual Manipulation: A Review
Inkyu Sa, Chanoh Park, Hea-Min Lee +2
Vision Language Action (VLA) models unify visual perception, natural-language understanding, and action generation within a single foundation model, allowing a robot to follow inst…
cs.RO2023★ 5 cited
A Sign Language Recognition System with Pepper, Lightweight-Transformer, and LLM
JongYoon Lim, Inkyu Sa, Bruce MacDonald +1
This research explores using lightweight deep neural network architectures to enable the humanoid robot Pepper to understand American Sign Language (ASL) and facilitate non-verbal…