3 papers
cs.AI2026
PatientAgentBench: A Benchmark Framework for Evaluating Patient-Facing Health AI Agents
Korosh Vatanparvar, Ashutosh Joshi, Maria Xenochristou +11
Health AI is evolving from answering questions to agentic systems that converse with patients, reason about health records, and act on their behalf. Primary care guards against dia…
cs.AI2026
IMCBench: A benchmark for multimodal LLMs in Image-grounded Medical Conversations
Maria Xenochristou, Ashutosh Joshi, Korosh Vatanparvar +10
Recent advances in large language models and vision-language models have enabled reasoning over multimodal data, offering opportunities for clinical applications such as decision s…
cs.CL2026
Linking Knowledge to Care: Knowledge Graph-Augmented Medical Follow-Up Question Generation
Liwen Sun, Xiang Yu, Ming Tan +4
Clinical diagnosis is time-consuming, requiring intensive interactions between patients and medical professionals. While large language models (LLMs) could ease the pre-diagnostic…