3 papers
cs.RO2026
Reducing Temporal Redundancy for Efficient Vision-Language-Action Inference
Yuzhou Wu, Yuxin Zheng, Muchun Niu +6
Vision-Language-Action (VLA) models exhibit strong generalization for robotic manipulation, yet their high inference latency limits real time deployment. We identify two primary so…
cs.RO2026
RoBoSR: Structured Scene Representations for Embodied Robotic Reasoning
Kewei Hu, Wanchan Yu, Fangwen Chen +6
Despite rapid progress, embodied reasoning under real-world variability remains challenging. Existing approaches rely on demonstration-driven sequential biases, limiting flexibilit…
cs.RO2026
GSR: Learning Structured Reasoning for Embodied Manipulation
Kewei Hu, Michael Zhang, Wei Ying +7
Despite rapid progress, embodied agents still struggle with long-horizon manipulation that requires maintaining spatial consistency, causal dependencies, and goal constraints. A ke…