2 papers
cs.RO2026
ARM: Advantage Reward Modeling for Long-Horizon Manipulation
Yiming Mao, Zixi Yu, Weixin Mao +5
Long-horizon robotic manipulation remains challenging for reinforcement learning (RL) because sparse rewards provide limited guidance for credit assignment. Practical policy improv…
astro-ph.IM2025
StarWhisper Telescope: An AI framework for automating end-to-end astronomical observations
Cunshi Wang, Yu Zhang, Yuyang Li +25
The exponential growth of large-scale telescope arrays has boosted time-domain astronomy development but introduced operational bottlenecks, including labor-intensive observation p…