1 paper
Yifan Zhang, Rahmatollah Beheshti
Automated detection of errors in clinical documentation is a promising application of large language models (LLMs), yet decisions to deploy such models rest on benchmarks that eval…