works on

From the 2 of 17 linked papers with an AI index.

most citedInvestigating the Dark Energy Constraint from Strongly Lensed AGN at LSST-Scale

2 citations · 2 across the 3 of their papers we have counts for

collaborators

17 papers

astro-ph.IM2026

Early Exploration of the Scientific Discovery Space for the Habitable Worlds Observatory

Courtney D. Dressing, Danica Adams, Evelyne Alecian +324

The paper summarizes 70 science cases for the proposed NASA Habitable Worlds Observatory, outlining the observational needs across four scientific pillars and detailing required ca…

astro-ph.CO20262 cited

Investigating the Dark Energy Constraint from Strongly Lensed AGN at LSST-Scale

Sydney Erickson, Martin Millon, Padmavathi Venkatraman +12

The paper presents a scalable hierarchical inference framework to jointly analyze hundreds of strongly lensed AGN time delays from LSST, forecasting a ~2.5% measurement of H0 and a…

cs.AI2026

OpenThoughts-Agent: Data Recipes for Agentic Models

Negin Raoof, Richard Zhuang, Marianna Nezhurina +47

Agentic language models dramatically expand the applications of AI yet little is publicly known about how to curate training data for broadly capable agents. Existing open efforts…

cs.AI2026

SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks

Xiangyi Li, Yimin Liu, Wenbo Chen +75

Agent Skills are structured packages of procedural knowledge that augment large language model (LLM) agents at inference time. Despite rapid adoption, there is no standard way to m…

cs.AI2026

Every Eval Ever: A Unifying Schema and Community Repository for AI Evaluation Results

Jan Batzner, Sree Harsha Nelaturu, Damian Stachura +45

AI evaluations are widely used for testing and understanding progress. However, the diverse evaluators bring with them inconsistencies that challenge analysis and comparison. First…

cs.SE2026

SWE-Marathon: Can Agents Autonomously Complete Ultra-Long-Horizon Software Work?

Rishi Desai, Jesse Hu, Joan Cabezas +23

AI agents are increasingly expected to complete long-horizon workflows that require sustained progress over hours, millions of tokens, and complex environments. Yet current agent b…