empirical reasoning 1formal verification 1hallucination mitigation 1large language models 1tool calling 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.LG2026
Evidence-Grounded Verified Agentic Reasoning: A Path Toward Eliminating LLM Hallucination in Empirical Inference via Tool-Attested Kernel Proofs
Junyu Ren
The paper presents EG-VAR, a system that integrates large language models with the Lean theorem prover so that every empirical claim is tied to an attested tool call and formally v…
cs.AI2026
Scaling Multiagent Systems with Process Rewards
Ed Li, Junyu Ren, Cat Yan
While multiagent systems have shown promise for tackling complex tasks via specialization, finetuning multiple agents simultaneously faces two key challenges: (1) credit assignment…
cs.AI2025
Build Your Personalized Research Group: A Multiagent Framework for Continual and Interactive Science Automation
Ed Li, Junyu Ren, Xintian Pan +4
The automation of scientific discovery represents a critical milestone in Artificial Intelligence (AI) research. However, existing agentic systems for science suffer from two funda…