3 papers
cs.AI2026
CEDAR-GRPO: Process-Aware Reinforcement Learning for General Abductive Reasoning in LLMs
Moein Salimi, Danial Parnian, Shaygan Adim +6
Abductive reasoning, often characterized as inference to the best explanation, is central to explanation under uncertainty, from everyday sense-making and investigation to scientif…
cs.AI2026
Wiring the 'Why': A Unified Taxonomy and Survey of Abductive Reasoning in LLMs
Moein Salimi, Shaygan Adim, Danial Parnian +3
Regardless of its foundational role in human discovery and sense-making, abductive reasoning--the inference of the most plausible explanation for an observation--has been relativel…
cs.AI2026
Debate as Reward: A Multi-Agent Reward System for Scientific Ideation via RL Post-Training
Moein Salimi, Babak Hosseini Mohtasham, Amin Aghakasiri +6
Large Language Models (LLMs) have demonstrated potential in automating scientific ideation, yet current approaches relying on iterative prompting or complex multi-agent architectur…