activity
20192026
most citedScaling Cross-Embodied Learning: One Policy for Manipulation, Navigation, Locomotion and Aviation

6 citations · 17 across the 24 of their papers we have counts for

collaborators
Showing cs.ROShow all

25 papers · 1 filter

cs.RO2026

Learning to Act While Waiting: RL Finetuning of Generalist Robot Policies Under Inference Latency

Brian Zhu, Momen Khalil, E Harrison +17

While reinforcement learning (RL) allows generalist robot policies to continually improve during deployment, the large model size of modern generalist policies, such as VLAs, poses…

cs.RO2026

FlowDAgger: Human-in-the-Loop Adaptation of Generative Robot Policies in Latent Space

Michael Murray, Daphne Chen, Simran Bagaria +7

Pretrained generative robot policies based on flow matching and diffusion have achieved impressive results across a wide range of manipulation tasks. Yet real-world deployments rou…

cs.RO2026

Robot Self-Improvement via Human-Video Dynamics Models

Hanzhi Chen, Anran Zhang, Simon Schaefer +5

A central question in robot learning is how to acquire skills from the kinds of data that humans learn from: passive observation, embodied practice, and the experience of failure.…

cs.RO2026

World Model for Robot Learning: A Comprehensive Survey

Bohan Hou, Gen Li, Jindou Jia +15

World models, which are predictive representations of how environments evolve under actions, have become a central component of robot learning. They support policy learning, planni…

cs.RO2026

SteerVLA: Steering Vision-Language-Action Models in Long-Tail Driving Scenarios

Tian Gao, Celine Tan, Catherine Glossop +8

A fundamental challenge in autonomous driving is the integration of high-level, semantic reasoning for long-tail events with low-level, reactive control for robust driving. While l…

cs.RO2025

mimic-video: Video-Action Models for Generalizable Robot Control Beyond VLAs

Jonas Pai, Liam Achenbach, Victoriano Montesinos +3

Prevailing Vision-Language-Action Models (VLAs) for robotic manipulation are built upon vision-language backbones pretrained on large-scale, but disconnected static web data. As a…