Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
ResearchClawBench: A Benchmark for End-to-End Autonomous Scientific Research
Wanghan Xu, Shuo Li, Tianlin Ye +48
AI coding agents are increasingly used for scientific work, but their end-to-end autonomous research capability remains difficult to verify. We present ResearchClawBench, a benchma…
cs.LG2025
A Self-Evolving AI Agent System for Climate Science
Zijie Guo, Jiong Wang, Fenghua Ling +19
Scientific progress in Earth science depends on integrating data across the planet's interconnected spheres. However, the accelerating volume and fragmentation of multi-sphere know…