5 citations · 5 across the 3 of their papers we have counts for
Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
SUP-MIMIC: A Multi-Task Clinical Diagnosis Benchmark for Evaluating LLMs' Robustness to Contradictory Evidence
Yi Yu, Bo Wang, Chong Feng +4
Current evaluations of large language models (LLMs) primarily focus on factual knowledge retrieval, overlooking the fundamental challenge of navigating the complex, non-bijective m…
cs.CL2026
Uncertainty-Aware Structured Data Extraction from Full CMR Reports via Distilled LLMs
Yi Yu, Parker Martin, Zhenyu Bu +5
Converting free-text cardiac magnetic resonance (CMR) reports into auditable structured data remains a bottleneck for cohort assembly, longitudinal curation, and clinical decision…
cs.CL2024★ 5 cited
A Survey for Large Language Models in Biomedicine
Chong Wang, Mengyao Li, Junjun He +14
Recent breakthroughs in large language models (LLMs) offer unprecedented natural language understanding and generation capabilities. However, existing surveys on LLMs in biomedicin…