activity
20242026
most citedTeamPath: Building MultiModal Pathology Experts with Reasoning AI Copilots

1 citations · 1 across the 4 of their papers we have counts for

collaborators
Showing cs.CLShow all

12 papers · 1 filter

cs.CL2026

Evaluating Multi-Turn Multimodal Diagnostic Reasoning on Challenging Real-World Clinical Cases

Rui Yang, Weihao Xuan, Yi Lin +23

Clinical diagnostic evaluation should not only assess whether models can provide correct diagnoses, but also reflect the realities of clinical practice, including progressive discl…

cs.CL2026

SIMAX: A Scalable and Interpretable Framework for Multi-Fidelity and Annotated Clinician-Patient Dialogue Simulation

Zhuhan Bao, Rui Yang, Bohao Yang +16

Background. The widespread deployment of ambient digital scribes is driving large-scale capture of clinician-patient dialogues. Human coding of clinical communication data remains…

cs.CL2026

Toward Global Large Language Models in Medicine

Rui Yang, Huitao Li, Weihao Xuan +47

Despite continuous advances in medical technology, the global distribution of health care resources remains uneven. The development of large language models (LLMs) has transformed…

cs.CL2025

Leveraging LLMs for Title and Abstract Screening for Systematic Review: A Cost-Effective Dynamic Few-Shot Learning Approach

Yun-Chung Liu, Rui Yang, Jonathan Chong Kai Liew +4

Systematic reviews are a key component of evidence-based medicine, playing a critical role in synthesizing existing research evidence and guiding clinical decisions. However, with…

cs.CL2025

An Agentic AI System for Multi-Framework Communication Coding

Bohao Yang, Rui Yang, Joshua M. Biro +14

Clinical communication is central to patient outcomes, yet large-scale human annotation of patient-provider conversation remains labor-intensive, inconsistent, and difficult to sca…

cs.CL2025

HealthContradict: Evaluating Biomedical Knowledge Conflicts in Language Models

Boya Zhang, Alban Bornet, Rui Yang +2

How do language models use contextual information to answer health questions? How are their responses impacted by conflicting contexts? We assess the ability of language models to…