2 papers
cs.AI2026
A global log for medical AI
Ayush Noori, Aaron E. Boussina, Hai Ho Bich +48
Modern computer systems rely on syslog, a universal protocol that records critical events across heterogeneous infrastructure. Medicine's rapidly growing AI stack has no equivalent…
cs.CL2024
A Proposed S.C.O.R.E. Evaluation Framework for Large Language Models : Safety, Consensus, Objectivity, Reproducibility and Explainability
Ting Fang Tan, Kabilan Elangovan, Jasmine Ong +10
A comprehensive qualitative evaluation framework for large language models (LLM) in healthcare that expands beyond traditional accuracy and quantitative metrics needed. We propose…