activity
20242026
collaborators
Showing cs.ROShow all

7 papers · 1 filter

cs.RO2026

Enhancing Visual Domain Robustness in Behaviour Cloning via Saliency-Guided Augmentation

Zheyu Zhuang, Ruiyu Wang, Nils Ingelhag +2

In vision-based behavior cloning (BC), conventional image augmentations such as Random Crop and Color Jitter often fall short under substantial visual domain shifts, including chan…

cs.RO2026

MirrorDuo: Reflection-Consistent Visuomotor Learning from Mirrored Demonstration Pairs

Zheyu Zhuang, Ruiyu Wang, Giovanni Luca Marchetti +2

Image-based behaviour cloning leverages demonstrations captured from ubiquitous RGB cameras. However, it remains constrained by the cost of collecting diverse demos, especially for…

cs.RO2026

PALM: Enhanced Generalizability for Local Visuomotor Policies via Perception Alignment

Ruiyu Wang, Zheyu Zhuang, Danica Kragic +1

Generalizing beyond the training domain in image-based behavior cloning remains challenging. Existing methods address individual axes of generalization, workspace shifts, viewpoint…

cs.RO2025

R900: Understanding the Cost-Effectiveness of Random Exploration from 900 Hours of Robotic Data Collection

Shutong Jin, Axel Kaliff, Ruiyu Wang +2

Data scarcity presents a key bottleneck for imitation learning in robotic manipulation. In this paper, we focus on random exploration data-actions and video sequences produced auto…

cs.RO2025

Feature Extractor or Decision Maker: Rethinking the Role of Visual Encoders in Visuomotor Policies

Ruiyu Wang, Zheyu Zhuang, Shutong Jin +3

An end-to-end (E2E) visuomotor policy is typically treated as a unified whole, but recent approaches using out-of-domain (OOD) data to pretrain the visual encoder have cleanly sepa…

cs.RO2024

PACA: Perspective-Aware Cross-Attention Representation for Zero-Shot Scene Rearrangement

Shutong Jin, Ruiyu Wang, Kuangyi Chen +1

Scene rearrangement, like table tidying, is a challenging task in robotic manipulation due to the complexity of predicting diverse object arrangements. Web-scale trained generative…