Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Verify, Repair, Repeat, or Stop? Robust Stopping for Noisy Verify-Repair Loops in LLM Agents
Yitao Wu, Si Shen, Rui Yang +2
Verify-repair loops are a standard means for large language model (LLM) agents to correct faulty plans in code generation, mathematical reasoning, and tool use. When both the verif…
cs.AI2026
LLM-Metrics: Measuring Research Impact Through Large Language Model Memory
Si Shen, Wenhua Zhao, Danhao Zhu
Citation counts remain the dominant metric for assessing research impact, yet they suffer from well-documented limitations: temporal lag, disciplinary bias, and Matthew effects. He…