Showing cs.ROShow all
2 papers · 1 filter
cs.RO2026
World Action Models Enable Continual Imitation Learning with Recurrent Generative Replays
Manish Kumar Govind, Dominick Reilly, Smit Patel +2
Going beyond predicting robot actions, World Action Models (WAMs) can also generate future visual observations. We build on this generative capability to propose Recurrent Generati…
cs.RO2026
UniLACT: Depth-Aware RGB Latent Action Learning for Vision-Language-Action Models
Manish Kumar Govind, Dominick Reilly, Pu Wang +1
Latent action representations learned from unlabeled videos have recently emerged as a promising paradigm for pretraining vision-language-action (VLA) models without explicit robot…