3 papers
cs.LG2026
Information-Time Proximal Policy Optimization
Yongcheng Zeng, Xinyu Cui, Yan Song +9
RLVR has substantially improved the reasoning capabilities of LLMs. However, existing methods typically parameterize temporal progression in the Markov Decision Process by token-by…
cs.RO2026
ME-Brain-1.0: Memory, Cognition and Action for Evolving Embodied Intelligence
Wei He, Hengtao Li, Zhongrui Yu +20
Current embodied systems largely rely on pretrained capabilities that remain fixed after deployment, limiting their ability to learn from physical interaction. We introduce MachEmb…
cs.CV2026
ME-Dex 1.0: Bringing Heterogeneous Tactile Sensing into World Action Modeling
Xuancheng Zhang, Xuetao Liu, Qianying Tang +8
World Action Models bring the predictive capabilities of video models into robot action generation, providing a rich foundation for modeling future visual states. Tactile sensing c…