4 papers · 1 filter
Beyond Individual Intelligence: Surveying Collaboration, Failure Attribution, and Self-Evolution in LLM-based Multi-Agent Systems
Shihao Qi, Jie Ma, Rui Xing +15
LLM-based autonomous agents have demonstrated strong capabilities in reasoning, planning, and tool use, yet remain limited when tasks require sustained coordination across roles, t…
ErrEval: Error-Aware Evaluation for Question Generation through Explicit Diagnostics
Weiping Fu, Bifan Wei, Jingyi Hao +7
Automatic Question Generation (QG) often produces outputs with critical defects, such as factual hallucinations and answer mismatches. However, existing evaluation methods, includi…
From Static to Dynamic: Adaptive Monte Carlo Search for Mathematical Process Supervision
Jie Ma, Shihao Qi, Rui Xing +4
The quality of process data plays a key role in training a Process Reward Model (PRM), which can enhance the complex mathematical reasoning capability of large language models. Exi…
GKG-LLM: A Unified Framework for Generalized Knowledge Graph Construction
Jian Zhang, Bifan Wei, Shihao Qi +3
The construction of Generalized Knowledge Graph (GKG), including knowledge graph, event knowledge graph and commonsense knowledge graph, is fundamental for various natural language…