2 papers
cs.AI2026
Making AI Evaluation Deployment Relevant Through Context Specification
Matthew Holmes, Thiago Lacerda, Reva Schwartz
With many organizations struggling to gain value from AI deployments, pressure to evaluate AI in an informed manner has intensified. Status quo AI evaluation approaches often mask…
cs.AI2026
CIRCLE: A Framework for Evaluating AI from a Real-World Lens
Reva Schwartz, Carina Westling, Morgan Briggs +12
This paper proposes CIRCLE, a six-stage, lifecycle-based framework to bridge the reality gap between model-centric performance metrics and AI's materialized outcomes in deployment.…