85 citations · 396 across the 47 of their papers we have counts for
3 papers · 1 filter
Asclepius: An Adaptive Harness for Long-Horizon Clinical Agents
Grace Chang Yuan, Xiaoman Zhang, Sung Eun Kim +2
LLM agents are predominantly benchmarked on short, single-task trajectories, yet real deployments run for hours under contention, surfacing a different class of failures. We use th…
HeadCT-ONE: Enabling Granular and Controllable Automated Evaluation of Head CT Radiology Report Generation
Julián N. Acosta, Xiaoman Zhang, Siddhant Dogra +5
We present Head CT Ontology Normalized Evaluation (HeadCT-ONE), a metric for evaluating head CT report generation through ontology-normalized entity and relation extraction. HeadCT…
Uncovering Knowledge Gaps in Radiology Report Generation Models through Knowledge Graphs
Xiaoman Zhang, Julián N. Acosta, Hong-Yu Zhou +1
Recent advancements in artificial intelligence have significantly improved the automatic generation of radiology reports. However, existing evaluation methods fail to reveal the mo…