2 papers
cs.HC2026
A prospective clinical feasibility study of a conversational diagnostic AI in an ambulatory primary care clinic
Peter Brodeur, Jacob M. Koshy, Anil Palepu +45
Large language model (LLM)-based AI systems have shown promise for patient-facing diagnostic and management conversations in simulated settings. Translating these systems into clin…
cs.LG2025
Evaluating AI systems under uncertain ground truth: a case study in dermatology
David Stutz, Ali Taylan Cemgil, Abhijit Guha Roy +17
For safety, medical AI systems undergo thorough evaluations before deployment, validating their predictions against a ground truth which is assumed to be fixed and certain. However…