2 papers
cs.CL2026
DA-RAC: Distance-Aware Calibration of LLM Judges for Trustworthy AI Auditing
Cheng Wu, Vishal Anand, Jaya Krishna Mandivarapu +2
Generative AI systems are increasingly producing real-world artifacts, however their efficacy and validity are often evaluated via context-free LLM-scoring. These judges can be mis…
cs.RO2026
AXIS: A Growable Community-Driven Data Engine for Scalable Robot Manipulation
Mengfei Zhao, Dihong Huang, Yikai Tang +12
Learning effective robot manipulation policies requires diverse, high-quality demonstrations, yet existing data pipelines are often difficult to scale because they rely on speciali…