3 papers
cs.LG2026
WeatherNext 3: Increasing resolution and performance of global weather models with raw observations
Stephan Rasp, Boris Babenko, Dominic Masters +22
State-of-the-art AI weather models have shown impressive medium-range forecast skill and computational efficiency, but suffer two key shortcomings: their forecasts have lower spati…
cs.AI2025
CLINB: A Climate Intelligence Benchmark for Foundational Models
Michelle Chen Huebscher, Katharine Mach, Aleksandar Stanić +10
Evaluating how Large Language Models (LLMs) handle complex, specialized knowledge remains a critical challenge. We address this through the lens of climate change by introducing CL…
cs.CL2023
Assessing Large Language Models on Climate Information
Jannis Bulian, Mike S. Schäfer, Afra Amini +8
As Large Language Models (LLMs) rise in popularity, it is necessary to assess their capability in critically relevant domains. We present a comprehensive evaluation framework, grou…