13 papers
RoboPIN: Grounded Embodied Reasoning via Pinned Chain-of-Thought
Yaoting Huang, Yifu Yuan, Linqi Han +6
Embodied reasoning requires models to perceive task-relevant objects and spaces in physical environments and maintain consistent visual grounding throughout multi-step reasoning. H…
AhaRobot: A Low-Cost Open-Source Bimanual Mobile Manipulator for Embodied AI
Haiqin Cui, Yifu Yuan, Yan Zheng +1
Scaling Vision-Language-Action models for embodied manipulation demands large volumes of diverse manipulation data, yet the high cost of commercial mobile manipulators and teleoper…
Embodied-R1: Reinforced Embodied Reasoning for General Robotic Manipulation
Yifu Yuan, Haiqin Cui, Yaoting Huang +7
Generalization in embodied AI is hindered by the "seeing-to-doing gap," which stems from data scarcity and embodiment heterogeneity. To address this, we pioneer "pointing" as a uni…
From Seeing to Doing: Bridging Reasoning and Decision for Robotic Manipulation
Yifu Yuan, Haiqin Cui, Yibin Chen +7
Achieving generalization in robotic manipulation remains a critical challenge, particularly for unseen scenarios and novel tasks. Current Vision-Language-Action (VLA) models, while…
LongCat-Flash-Thinking-2601 Technical Report
Meituan LongCat Team, Anchun Gui, Bei Li +162
We introduce LongCat-Flash-Thinking-2601, a 560-billion-parameter open-source Mixture-of-Experts (MoE) reasoning model with superior agentic reasoning capability. LongCat-Flash-Thi…
Key Decision-Makers in Multi-Agent Debates: Who Holds the Power?
Qian Zhang, Yan Zheng, Jinyi Liu +2
Recent studies on LLM agent scaling have highlighted the potential of Multi-Agent Debate (MAD) to enhance reasoning abilities. However, the critical aspect of role allocation strat…