15 papers
DemoBridge: A Simulation-in-the-Loop Toolkit for Single-View Human Demonstration Retargeting
Zehao Wang, Fabien Despinoy, Sergey Zakharov +2
We present DemoBridge, an toolkit that turns a single-view RGB stereo recording of a human hand demonstration into an executable, physics-validated robot-arm trajectory. Retargetin…
Agents' Last Exam
Yiyou Sun, Xinyang Han, Weichen Zhang +306
Recent AI systems have achieved strong results on a wide range of benchmarks, yet these gains have not translated into economically meaningful deployment across many professional d…
TAGA: Terrain-aware Active Gaze Learning for Generalizable Agile Humanoid Locomotion
Peizhuo Li, Hongyi Li, Mingfeng Fan +9
Agile humanoid locomotion across diverse challenging terrain demands both wide perceptual coverage and precise local geometry understanding. Motivated by the way humans selectively…
HMPO: Hybrid Median-length Policy Optimization for Chain-of-Thought Compression
Minghui Zheng, Hongxu Chen, Huimin Ren +8
Large language models achieve remarkable performance via extended chain-of-thought (CoT) reasoning, yet this lengthy process incurs substantial inference overhead. Existing CoT com…
Large Language Models in Transportation Systems Management and Operations: From Text Reasoning to Multi-modal Decision Support
Siyan Li, Zehao Wang, Jiachen Li +3
Transportation systems management and operations (TSMO) increasingly depends on timely interpretation of heterogeneous data, from various sensor streams, incident reports, traveler…
From Demonstrations to Rewards: Test-Time Prompt Optimization for VLM Reward Models
Christian Gumbsch, Leonardo Barcellona, Lennard Schünemann +7
Reinforcement learning relies on accurate reward functions, which are often hand-crafted or even unavailable in real-world applications, such as robotics. Recent work has explored…