3 papers
cs.AI2026
AstaBench: Rigorous Benchmarking of AI Agents with a Scientific Research Suite
Jonathan Bragg, Mike D'Arcy, Nishant Balepur +36
AI agents hold the potential to revolutionize scientific productivity by automating literature reviews, replicating experiments, analyzing data, and even proposing new directions o…
cs.CL2026
Generating Literature-Driven Scientific Theories at Scale
Peter Jansen, Peter Clark, Doug Downey +1
Contemporary automated scientific discovery has focused on agents for generating scientific experiments, while systems that perform higher-level scientific activities such as theor…
cs.AI2025
HARPA: A Testability-Driven, Literature-Grounded Framework for Research Ideation
Rosni Vasu, Peter Jansen, Pao Siangliulue +4
While there has been a surge of interest in automated scientific discovery (ASD), especially with the emergence of LLMs, it remains challenging for tools to generate hypotheses tha…