2 papers
cs.CL2026
Clinician use of language models diverges from how the models are evaluated
Krithik Vishwanath, Haitong Lin, Anton Alyakin +19
Large language model (LLM) assistants are being deployed to clinicians across health systems, and judgments about their readiness rest largely on benchmark scores, most of them der…
cs.CL2026
Large language models are vulnerable to incidental information in clinical documentation and reasoning
Krithik Vishwanath, Brandon Ye, Anton Alyakin +4
Large language models (LLMs) are increasingly relied upon to support ambient documentation and clinical reasoning. Here we examine the impact of a failure mode shared between these…