papers

Publications (13)

cs.RO2024

Fast Explicit-Input Assistance for Teleoperation in Clutter

Nick Walker, Xuning Yang, Animesh Garg +3

The performance of prediction-based assistance for robot teleoperation degrades in unseen or goal-rich environments due to incorrect or quickly-changing intent inferences. Poor pre…

cs.RO2025

Robot Policy Evaluation for Sim-to-Real Transfer: A Benchmarking Perspective

Xuning Yang, Clemens Eppner, Jonathan Tremblay +3

Current vision-based robotics simulation benchmarks have significantly advanced robotic manipulation research. However, robotics is fundamentally a real-world problem, and evaluati…

cs.RO2019

Fast and Agile Vision-Based Flight with Teleoperation and Collision Avoidance on a Multirotor

Alex Spitzer, Xuning Yang, John Yao +6

We present a multirotor architecture capable of aggressive autonomous flight and collision-free teleoperation in unstructured, GPS-denied environments. The proposed system enables…

cs.RO2025

Inference-Time Policy Steering through Human Interactions

Yanwei Wang, Lirui Wang, Yilun Du +6

Generative policies trained with human demonstrations can autonomously accomplish multimodal, long-horizon tasks. However, during inference, humans are often removed from the polic…

cs.RO2026

Learning to Plan & Schedule with Reinforcement-Learned Bimanual Robot Skills

Weikang Wan, Fabio Ramos, Xuning Yang +1

Long-horizon contact-rich bimanual manipulation presents a significant challenge, requiring complex coordination involving a mixture of parallel execution and sequential collaborat…

cs.CV2026

Adaptive Volumetric Mechanical Property Fields Invariant to Resolution

Rishit Dagli, Donglai Xiang, Vismay Modi +4

Accurate mechanical properties (or materials) Young's modulus (), Poisson's ratio () and density () are essential for reliable physics simulation of digital worlds, but…

cs.CV2026

Cosmos 3: Omnimodal World Models for Physical AI

NVIDIA, :, Aditi +293

We introduce Cosmos 3, a family of omnimodal world models designed to jointly process and generate language, image, video, audio, and action sequences within a unified mixture-of-t…

cs.RO2025

VLA-0: Building State-of-the-Art VLAs with Zero Modification

Ankit Goyal, Hugo Hadfield, Xuning Yang +2

Vision-Language-Action models (VLAs) hold immense promise for enabling generalist robot manipulation. However, the best way to build them remains an open question. Current approach…

cs.RO2026

VoLo: A Physical Orchestrator for Open-Vocabulary Long-Horizon Manipulation

Siyi Chen, Hugo Hadfield, Alex Zook +9

Open-vocabulary long-horizon manipulation requires robots to reason over flexible instructions and complex multi-object scenes while adaptively planning, executing, monitoring, and…

cs.RO2026

RoboLab: A High-Fidelity Simulation Benchmark for Analysis of Task Generalist Policies

Xuning Yang, Rishit Dagli, Alex Zook +5

The pursuit of general-purpose robotics has yielded impressive foundation models, yet simulation-based benchmarking remains a bottleneck due to rapid performance saturation and a l…

cs.RO2024

Aim My Robot: Precision Local Navigation to Any Object

Xiangyun Meng, Xuning Yang, Sanghun Jung +4

Existing navigation systems mostly consider "success" when the robot reaches within 1m radius to a goal. This precision is insufficient for emerging applications where the robot ne…

cs.RO2025

RoboArena: Distributed Real-World Evaluation of Generalist Robot Policies

Pranav Atreya, Karl Pertsch, Tony Lee +29

Comprehensive, unbiased, and comparable evaluation of modern generalist policies is uniquely challenging: existing approaches for robot benchmarking typically rely on heavy standar…

cs.RO2022

An imminent collision monitoring system with safe stopping interventions for autonomous aerial flights

Jasmine Cheng, Xuning Yang, Nathan Michael

Collision avoidance requires tradeoffs in planning time horizons. Depending on the planner, safety cannot always be guaranteed in uncertain environments given map updates. To mitig…