2 papers
cs.RO2026
RoboEval: Where Robotic Manipulation Meets Structured and Scalable Evaluation
Yi Ru Wang, Carter Ung, Christopher Tan +11
We introduce RoboEval, a structured evaluation framework and benchmark for robotic manipulation that augments binary success with principled behavioral and outcome metrics. Existin…
cs.RO2026
RoboPlayground: Democratizing Robotic Evaluation through Structured Physical Domains
Yi Ru Wang, Carter Ung, Evan Gubarev +3
Evaluation of robotic manipulation systems has largely relied on fixed benchmarks authored by a small number of experts, where task instances, constraints, and success criteria are…