Showing cs.ROShow all
3 papers · 1 filter
cs.RO2026
Steering Autoregressive Vision-Language-Action Policies via Action Token Intervention
Jason Chan, Jonathan C. Kao
We present Token Steering (TS), a method for dynamically steering trajectories generated by an autoregressive vision-language-action (VLA) model through direct intervention in the…
cs.RO2026
Flow Control: Steering Vision-Language-Action Models with Simple Real-Time Inputs
Jonathan C. Kao, Jason Chan, Andy Wang
We introduce flow control of vision-language-action (VLA) models, a simple and effective way to steer VLA actions in real-time through generic inputs, such as a keyboard. This meth…
cs.RO2026
Intent at a Glance: Gaze-Guided Robotic Manipulation via Foundation Models
Tracey Yee Hsin Tay, Xu Yan, Jonathan Ouyang +4
Designing intuitive interfaces for robotic control remains a central challenge in enabling effective human-robot interaction, particularly in assistive care settings. Eye gaze offe…