3 papers
cs.RO2026
Primitive Subspaces Mediate Few-Shot Transfer in VLAs
Anya Singh, Cabrel Happi, Jai Relan +2
Deploying vision-language-action (VLA) policies in industrial environments requires the ability to teach new tasks at low cost, a property current VLAs lack, since each new task re…
cs.CV2026
WristCompass: Kinematic Coupling as a Learnable Visual Concept for Ego-Camera Orientation
Varun Nair, Vidyut Baradwaj, Jiahang He +3
Recovering ego-camera orientation from manipulation video is a prerequisite for disentangling hand motion from camera motion, a key step in imitation learning from egocentric demon…
cs.LG2026
BOKBO (Best of K Bad Options): Calibrated Abstention for VLA Policies
Anya Singh, Cabrel Happi, Jai Relan +2
Test-time scaling for vision-language-action (VLA) policies, methods such as RoboMonkey, SEAL, MG-Select, and V-GPS, samples K candidate action chunks at inference and executes the…