#multi-step reasoning
topicmulti-step reasoning
2 papers · 1 filter
cs.LG2026
ClawTrack: Towards Trace-Level Evaluation and Improvement of Real-World Autonomous Agents
Xingjian Wu, Xuhang Zhu, Xingchen Liu +6
The paper introduces ClawTrack, a benchmark that evaluates both the final outcomes and the step-by-step reasoning processes of LLM-based autonomous agents across multiple dimension…
cs.AI2026
DeepResearch Agent System
Yong Huang, Yulu Huang, for the team Collaboration
The DeepResearch Agent System is a large language model designed for deep information retrieval and multi-step autonomous research, using a sparse activation architecture that acti…