1 citations · 1 across the 2 of their papers we have counts for
3 papers
cs.CL2025
AlphaOne: Reasoning Models Thinking Slow and Fast at Test Time
Junyu Zhang, Runpei Dong, Han Wang +8
This paper presents AlphaOne (1), a universal framework for modulating reasoning progress in large reasoning models (LRMs) at test time. 1 first introduces moment, which…
cs.RO2025
Learning Getting-Up Policies for Real-World Humanoid Robots
Xialin He, Runpei Dong, Zixuan Chen +1
Automatic fall recovery is a crucial prerequisite before humanoid robots can be reliably deployed. Hand-designing controllers for getting up is difficult because of the varied conf…
cs.RO2024★ 1 cited
Learning Smooth Humanoid Locomotion through Lipschitz-Constrained Policies
Zixuan Chen, Xialin He, Yen-Jen Wang +8
Reinforcement learning combined with sim-to-real transfer offers a general framework for developing locomotion controllers for legged robots. To facilitate successful deployment in…