From the 2 of 17 linked papers with an AI index.
2 citations · 2 across the 5 of their papers we have counts for
17 papers
Early Exploration of the Scientific Discovery Space for the Habitable Worlds Observatory
Courtney D. Dressing, Danica Adams, Evelyne Alecian +324
The paper summarizes 70 science cases for the proposed NASA Habitable Worlds Observatory, outlining the observational needs across four scientific pillars and detailing required ca…
Investigating the Dark Energy Constraint from Strongly Lensed AGN at LSST-Scale
Sydney Erickson, Martin Millon, Padmavathi Venkatraman +12
The paper presents a scalable hierarchical inference framework to jointly analyze hundreds of strongly lensed AGN time delays from LSST, forecasting a ~2.5% measurement of H0 and a…
OpenThoughts-Agent: Data Recipes for Agentic Models
Negin Raoof, Richard Zhuang, Marianna Nezhurina +47
Agentic language models dramatically expand the applications of AI yet little is publicly known about how to curate training data for broadly capable agents. Existing open efforts…
SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks
Xiangyi Li, Yimin Liu, Wenbo Chen +75
Agent Skills are structured packages of procedural knowledge that augment large language model (LLM) agents at inference time. Despite rapid adoption, there is no standard way to m…
Every Eval Ever: A Unifying Schema and Community Repository for AI Evaluation Results
Jan Batzner, Sree Harsha Nelaturu, Damian Stachura +45
AI evaluations are widely used for testing and understanding progress. However, the diverse evaluators bring with them inconsistencies that challenge analysis and comparison. First…
SWE-Marathon: Can Agents Autonomously Complete Ultra-Long-Horizon Software Work?
Rishi Desai, Jesse Hu, Joan Cabezas +23
AI agents are increasingly expected to complete long-horizon workflows that require sustained progress over hours, millions of tokens, and complex environments. Yet current agent b…