activity
20242026
collaborators

10 papers

cs.AI2026

EMAS: Stabilizing Multi-Agent System Evolution through Evidence-Guided Revision

Chao Fei, Qingyi Si, Kaihua Liang +3

Many methods for automated multi-agent system design optimize prompts and topologies during an initial design stage and then deploy the resulting system unchanged on subsequent sam…

cs.CL2026

Outcome-Grounded Advantage Reshaping for Fine-Grained Credit Assignment in Mathematical Reasoning

Ziheng Li, Liu Kang, Feng Xiao +7

Group Relative Policy Optimization (GRPO) has emerged as a promising critic-free reinforcement learning paradigm for reasoning tasks. However, standard GRPO employs a coarse-graine…

cs.AI2026

When Agents Evolve, Institutions Follow

Chao Fei, Hongcheng Guo, Yanghua Xiao

Across millennia, complex societies have faced the same coordination problem of how to organize collective action among cognitively bounded and informationally incomplete individua…

cs.LG2026

On the Trainability of Masked Diffusion Language Models via Blockwise Locality

Yuxiang Wang, Yu Xiang, Baojian Zhou +4

Masked diffusion language models (MDMs) have recently emerged as a promising alternative to standard autoregressive large language models (AR-LLMs), yet their optimization can be s…

cs.AI2026

Logics-STEM: Empowering LLM Reasoning via Failure-Driven Post-Training and Document Knowledge Enhancement

Mingyu Xu, Cheng Fang, Keyue Jiang +16

We present Logics-STEM, a state-of-the-art reasoning model fine-tuned on Logics-STEM-SFT-Dataset, a high-quality and diverse dataset at 10M scale that represents one of the largest…

cs.LG2025

Accelerated Evolving Set Processes for Local PageRank Computation

Binbin Huang, Luo Luo, Yanghua Xiao +2

This work proposes a novel framework based on nested evolving set processes to accelerate Personalized PageRank (PPR) computation. At each stage of the process, we employ a localiz…