12 citations · 27 across the 13 of their papers we have counts for
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
Coalition-Aware Skill Reliability for Self-Evolving Agents
Qiyan Zhao, Xiaofeng Zhang, Bo Liu +11
Agent skills, structured artifacts distilled from interaction trajectories and dynamically reused from skill banks, have become a central mechanism for enabling large language mode…
cs.AI2025
StepWiser: Stepwise Generative Judges for Wiser Reasoning
Wei Xiong, Wenting Zhao, Weizhe Yuan +4
As models increasingly leverage multi-step reasoning strategies to solve complex problems, supervising the logical validity of these intermediate steps has become a critical resear…
cs.AI2025
Self-rewarding correction for mathematical reasoning
Wei Xiong, Hanning Zhang, Chenlu Ye +3
We study self-rewarding reasoning large language models (LLMs), which can simultaneously generate step-by-step reasoning and evaluate the correctness of their outputs during the in…