2 papers
cs.RO2026
LIBERO-VPro: Benchmarking Closed-Loop Visual Robustness of Robotic Foundation Models
Huiqiong Li, Zhiting Mei, Anirudha Majumdar +3
Robotic foundation models achieve impressive performance on standard manipulation benchmarks, yet these evaluations typically assume clean, timely, and consistent visual observatio…
cs.CV2026
RoboTrustBench: Benchmarking the Trustworthiness of Video World Models for Robotic Manipulation
Huiqiong Li, Jiayu Wang, Zhiting Mei +3
Video world models are increasingly used in robotic manipulation, yet existing benchmarks mostly evaluate them under valid, feasible, and safe instructions. We introduce RoboTrustB…