Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Loud or Silent? A Reusable Framework for Per-Modality Failure Analysis in Multimodal Clinical AI
Quang Bui, Shlok Jaiswal, Samuel Paik-Heintz +14
Multimodal clinical models are usually judged on accuracy with every modality present, but deployment removes modalities; an echocardiogram is often unavailable where an ECG is rou…
cs.AI2026
Testing the Black Box: Structural Barriers to Independent Evaluation of Consumer-Facing Health LLMs
Rahul Gorijavolu, Kaushik Madapati, Pritika Vig +7
Background: Consumer-facing large language models are now a common source of health information, and they interpret and personalize responses rather than retrieve them. Whether the…