3 papers
cs.RO2026
An Empirical Study and Open Testbed for Federated Fine-Tuning of Vision-Language-Action Models
Zhekai Duan, Kevin Ziyang Xie, Xinyu Tan +5
Adapting a pretrained Vision-Language-Action (VLA) model to a new robot, environment, or task requires demonstrations that are collected locally and often discarded. Federated lear…
cs.RO2026
Seeing is Not Believing: Breaking the Physical-to-Digital Trust Boundary in Robotics
Leming Shen, Shikai Geng, Yuanqing Zheng +1
In multi-robot collaboration, task handovers rely on downstream verifiers performing remote attestation, which inspects sensor telemetry to ensure a robot's physical behavior stric…
cs.RO2025
Fast ECoT: Efficient Embodied Chain-of-Thought via Thoughts Reuse
Zhekai Duan, Yuan Zhang, Shikai Geng +3
Embodied Chain-of-Thought (ECoT) reasoning enhances vision-language-action (VLA) models by improving performance and interpretability through intermediate reasoning steps. However,…