activity
20242026
most citedAstaBench: Rigorous Benchmarking of AI Agents with a Scientific Research Suite

1 citations · 1 across the 6 of their papers we have counts for

collaborators
Showing 2025Show all

5 papers · 1 filter

cs.CL2025

Ai2 Scholar QA: Organized Literature Synthesis with Attribution

Amanpreet Singh, Joseph Chee Chang, Chloe Anastasiades +15

Retrieval-augmented generation is increasingly effective in answering scientific questions from literature, but many state-of-the-art systems are expensive and closed-source. We in…

cs.IR2025

Literature-Grounded Novelty Assessment of Scientific Ideas

Simra Shahid, Marissa Radensky, Raymond Fok +3

Automated scientific idea generation systems have made remarkable progress, yet the automatic evaluation of idea novelty remains a critical and underexplored challenge. Manual eval…

cs.HC2025

Facets, Taxonomies, and Syntheses: Navigating Structured Representations in LLM-Assisted Literature Review

Raymond Fok, Joseph Chee Chang, Marissa Radensky +4

Comprehensive literature review requires synthesizing vast amounts of research -- a labor intensive and cognitively demanding process. Most prior work focuses either on helping res…

cs.AI2025

CodeScientist: End-to-End Semi-Automated Scientific Discovery with Code-based Experimentation

Peter Jansen, Oyvind Tafjord, Marissa Radensky +6

Despite the surge of interest in autonomous scientific discovery (ASD) of software artifacts (e.g., improved ML algorithms), current ASD systems face two key limitations: (1) they…

cs.HC2025

Toward Living Narrative Reviews: An Empirical Study of the Processes and Challenges in Updating Survey Articles in Computing Research

Raymond Fok, Alexa Siu, Daniel S. Weld

Surveying prior literature to establish a foundation for new knowledge is essential for scholarly progress. However, survey articles are resource-intensive and challenging to creat…