works on

From the 1 of 10 linked papers with an AI index.

activity
20242026
most citedOmniSpatial: Towards Comprehensive Spatial Reasoning Benchmark for Vision Language Models

1 citations · 1 across the 6 of their papers we have counts for

collaborators
Showing 2026Show all

5 papers · 1 filter

cs.RO2026

HumanTracker: Towards Comprehensive and Human-Aligned Motion Tracking Benchmark

Dairu Liu, Zekun Qi, Jiayu Zeng +11

Humanoid motion tracking is central to teleoperation and whole-body imitation, yet evaluation often disagrees with what people perceive in videos. Kinematic errors average per-fram…

cs.RO2026

Scaling Behavior Foundation Model for Humanoid Robots

Weishuai Zeng, Kangning Yin, Xiaojie Niu +15

The paper proposes a scalable behavior foundation model for humanoid robots that uses a motion‑tracking learning paradigm, coordinated on‑policy rollouts and diverse reference moti…

cs.RO2026

Humanoid-GPT: Scaling Data and Structure for Zero-Shot Motion Tracking

Zekun Qi, Xuchuan Chen, Dairu Liu +10

We introduce Humanoid-GPT, a GPT-style Transformer with causal attention trained on a billion-scale motion corpus for whole-body control. Unlike prior shallow MLP trackers constrai…

cs.RO2026

Seed2Scale: A Self-Evolving Data Engine for Embodied AI via Small to Large Model Synergy and Multimodal Evaluation

Cong Tai, Zhaoyu Zheng, Haixu Long +12

Existing data generation methods suffer from exploration limits, embodiment gaps, and low signal-to-noise ratios, leading to performance degradation during self-iteration. To addre…

cs.CV20261 cited

OmniSpatial: Towards Comprehensive Spatial Reasoning Benchmark for Vision Language Models

Mengdi Jia, Zekun Qi, Shaochen Zhang +5

Spatial reasoning is a key aspect of cognitive psychology and remains a bottleneck for current vision-language models (VLMs). While extensive research has aimed to evaluate or impr…