4 papers · 1 filter
MedRoundsQA: A Persona and Difficulty Aware Evaluation for Multi-Turn Medical Consultations
Youssef Mohamed, Ahmed Heakl, Qinrong Cui +13
Medical benchmarks are dominated by single-turn, multiple-choice clinical cases that poorly reflect real consultations. Practically, clinicians elicit evidence interactively and pa…
MAGE: Multi-Agent Self-Evolution with Co-Evolutionary Knowledge Graphs
Ruiyi Yang, Zechen Li, Hao Xue +2
Self-evolving language-model agents must decide what to learn next and how to preserve what they have learned across iterations. Existing systems typically carry this cross-iterati…
DoAtlas-1: A Causal Compilation Paradigm for Clinical AI
Yulong Li, Jianxu Chen, Xiwei Liu +9
Medical foundation models generate narrative explanations but cannot quantify intervention effects, detect evidence conflicts, or validate literature claims, limiting clinical audi…
Beyond Single Pass, Looping Through Time: KG-IRAG with Iterative Knowledge Retrieval
Ruiyi Yang, Hao Xue, Imran Razzak +2
Graph Retrieval-Augmented Generation (GraphRAG) has proven highly effective in enhancing the performance of Large Language Models (LLMs) on tasks that require external knowledge. B…