adversarial attacks 1automated security evaluation 1clinical AI auditing 1healthcare datasets 1large language models 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.CR2026
Evaluating Frontier AI Agents as Autonomous Clinical Security Auditors
Michael O. Eniolade
The paper introduces an evaluation framework where frontier AI agents autonomously perform security audits on clinical prediction models by executing adversarial attacks, computing…
cs.LG2026
StepShield: When, Not Whether to Intervene on Rogue Agents
Gloria Felicia, Zitha Sasindran, Jinfeng He +3
Agent safety benchmarks measure whether a monitor detects harm, not when. Yet timing is the difference between intervention and autopsy. We introduce StepShield, the first benchmar…
cs.LG2026
Calibration, Uncertainty Communication, and Deployment Readiness in CKD Risk Prediction: A Framework Evaluation Study
Michael O. Eniolade
Machine learning models for chronic kidney disease (CKD) risk prediction often post strong discrimination scores on internal test sets. Calibration and uncertainty quantification g…