Publications (13)
Fast Explicit-Input Assistance for Teleoperation in Clutter
Nick Walker, Xuning Yang, Animesh Garg +3
The performance of prediction-based assistance for robot teleoperation degrades in unseen or goal-rich environments due to incorrect or quickly-changing intent inferences. Poor pre…
Robot Policy Evaluation for Sim-to-Real Transfer: A Benchmarking Perspective
Xuning Yang, Clemens Eppner, Jonathan Tremblay +3
Current vision-based robotics simulation benchmarks have significantly advanced robotic manipulation research. However, robotics is fundamentally a real-world problem, and evaluati…
Fast and Agile Vision-Based Flight with Teleoperation and Collision Avoidance on a Multirotor
Alex Spitzer, Xuning Yang, John Yao +6
We present a multirotor architecture capable of aggressive autonomous flight and collision-free teleoperation in unstructured, GPS-denied environments. The proposed system enables…
Inference-Time Policy Steering through Human Interactions
Yanwei Wang, Lirui Wang, Yilun Du +6
Generative policies trained with human demonstrations can autonomously accomplish multimodal, long-horizon tasks. However, during inference, humans are often removed from the polic…
Learning to Plan & Schedule with Reinforcement-Learned Bimanual Robot Skills
Weikang Wan, Fabio Ramos, Xuning Yang +1
Long-horizon contact-rich bimanual manipulation presents a significant challenge, requiring complex coordination involving a mixture of parallel execution and sequential collaborat…
Adaptive Volumetric Mechanical Property Fields Invariant to Resolution
Rishit Dagli, Donglai Xiang, Vismay Modi +4
Accurate mechanical properties (or materials) Young's modulus (), Poisson's ratio () and density () are essential for reliable physics simulation of digital worlds, but…
Cosmos 3: Omnimodal World Models for Physical AI
NVIDIA, :, Aditi +293
We introduce Cosmos 3, a family of omnimodal world models designed to jointly process and generate language, image, video, audio, and action sequences within a unified mixture-of-t…
VLA-0: Building State-of-the-Art VLAs with Zero Modification
Ankit Goyal, Hugo Hadfield, Xuning Yang +2
Vision-Language-Action models (VLAs) hold immense promise for enabling generalist robot manipulation. However, the best way to build them remains an open question. Current approach…
VoLo: A Physical Orchestrator for Open-Vocabulary Long-Horizon Manipulation
Siyi Chen, Hugo Hadfield, Alex Zook +9
Open-vocabulary long-horizon manipulation requires robots to reason over flexible instructions and complex multi-object scenes while adaptively planning, executing, monitoring, and…
RoboLab: A High-Fidelity Simulation Benchmark for Analysis of Task Generalist Policies
Xuning Yang, Rishit Dagli, Alex Zook +5
The pursuit of general-purpose robotics has yielded impressive foundation models, yet simulation-based benchmarking remains a bottleneck due to rapid performance saturation and a l…
Aim My Robot: Precision Local Navigation to Any Object
Xiangyun Meng, Xuning Yang, Sanghun Jung +4
Existing navigation systems mostly consider "success" when the robot reaches within 1m radius to a goal. This precision is insufficient for emerging applications where the robot ne…
RoboArena: Distributed Real-World Evaluation of Generalist Robot Policies
Pranav Atreya, Karl Pertsch, Tony Lee +29
Comprehensive, unbiased, and comparable evaluation of modern generalist policies is uniquely challenging: existing approaches for robot benchmarking typically rely on heavy standar…
An imminent collision monitoring system with safe stopping interventions for autonomous aerial flights
Jasmine Cheng, Xuning Yang, Nathan Michael
Collision avoidance requires tradeoffs in planning time horizons. Depending on the planner, safety cannot always be guaranteed in uncertain environments given map updates. To mitig…