4 papers · 1 filter
CLExEval: A Human-in-the-Loop Framework for Qualitative Evaluation of LLM Clinical Reasoning
Ajmal M., Abin Roy, Afthab Salam Kanniyan +4
Large Language Models (LLMs) achieve strong results on many medical benchmarks, but their clinical reasoning remains difficult to evaluate reliably. A central risk is an evaluation…
InformGen: An AI Copilot for Accurate and Compliant Clinical Research Consent Document Generation
Zifeng Wang, Junyi Gao, Benjamin Danek +5
Leveraging large language models (LLMs) to generate high-stakes documents, such as informed consent forms (ICFs), remains a significant challenge due to the extreme need for regula…
A Perspective for Adapting Generalist AI to Specialized Medical AI Applications and Their Challenges
Zifeng Wang, Hanyin Wang, Benjamin Danek +6
The integration of Large Language Models (LLMs) into medical applications has sparked widespread interest across the healthcare industry, from drug discovery and development to cli…
Accelerating Clinical Evidence Synthesis with Large Language Models
Zifeng Wang, Lang Cao, Benjamin Danek +3
Synthesizing clinical evidence largely relies on systematic reviews of clinical trials and retrospective analyses from medical literature. However, the rapid expansion of publicati…