7 citations · 15 across the 22 of their papers we have counts for
10 papers · 1 filter
RoboPIN: Grounded Embodied Reasoning via Pinned Chain-of-Thought
Yaoting Huang, Yifu Yuan, Linqi Han +6
Embodied reasoning requires models to perceive task-relevant objects and spaces in physical environments and maintain consistent visual grounding throughout multi-step reasoning. H…
LongCat-Flash-Thinking-2601 Technical Report
Meituan LongCat Team, Anchun Gui, Bei Li +162
We introduce LongCat-Flash-Thinking-2601, a 560-billion-parameter open-source Mixture-of-Experts (MoE) reasoning model with superior agentic reasoning capability. LongCat-Flash-Thi…
Key Decision-Makers in Multi-Agent Debates: Who Holds the Power?
Qian Zhang, Yan Zheng, Jinyi Liu +2
Recent studies on LLM agent scaling have highlighted the potential of Multi-Agent Debate (MAD) to enhance reasoning abilities. However, the critical aspect of role allocation strat…
Reflex First, Reflect Later: Latency-Aware Embodied LLM Agents for Dynamic Response
Yangqing Zheng, Shunqi Mao, Dingxin Zhang +1
Large language models (LLMs) have substantially improved the planning capabilities of embodied agents, enabling their deployment in dynamic and safety-critical environments. Howeve…
CellAgent: An LLM-driven Multi-Agent Framework for Automated Single-cell Data Analysis
Yihang Xiao, Jinyi Liu, Yan Zheng +9
Single-cell RNA sequencing (scRNA-seq) data analysis is crucial for biological research, as it enables the precise characterization of cellular heterogeneity. However, manual manip…
MFE-ETP: A Comprehensive Evaluation Benchmark for Multi-modal Foundation Models on Embodied Task Planning
Min Zhang, Xian Fu, Jianye Hao +5
In recent years, Multi-modal Foundation Models (MFMs) and Embodied Artificial Intelligence (EAI) have been advancing side by side at an unprecedented pace. The integration of the t…