11 citations · 14 across the 2 of their papers we have counts for
3 papers · 1 filter
MEDEC: A Benchmark for Medical Error Detection and Correction in Clinical Notes
Asma Ben Abacha, Wen-wai Yim, Yujuan Fu +4
Several studies showed that Large Language Models (LLMs) can answer medical questions correctly, even outperforming the average human score in some medical exams. However, to our k…
ACI-BENCH: a Novel Ambient Clinical Intelligence Dataset for Benchmarking Automatic Visit Note Generation
Wen-wai Yim, Yujuan Fu, Asma Ben Abacha +3
Recent immense breakthroughs in generative models such as in GPT4 have precipitated re-imagined ubiquitous usage of these models in all applications. One area that can benefit by i…
An Investigation of Evaluation Metrics for Automated Medical Note Generation
Asma Ben Abacha, Wen-wai Yim, George Michalopoulos +1
Recent studies on automatic note generation have shown that doctors can save significant amounts of time when using automatic clinical note generation (Knoll et al., 2022). Summari…