2 papers
cs.AI2026
End-to-End Evaluation and Governance of an EHR-Embedded AI Agent for Clinicians
Aaryan Shah, Andrew Hines, Alexia Downs +6
Clinical AI systems require not just point-in-time evaluation but continuous governance: the ongoing practice of monitoring, evaluating, iterating, and re-evaluating performance th…
cs.AI2026
Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters
Aaryan Shah, Andrew Hines, Alexia Downs +6
Objective. Clinical AI documentation systems require evaluation methodologies that are clinically valid, economically viable, and sensitive to iterative changes. Methods requiring…