4 papers
Scalable Stewardship of an LLM-Assisted Clinical Benchmark with Physician Oversight
Junze Ye, Daniel Tawfik, Alex J. Goodell +3
Reference labels for machine-learning benchmarks are increasingly synthesized with LLM assistance, but their reliability remains underexamined. We audit MedCalc-Bench, a clinical b…
On Aligning Prediction Models with Clinical Experiential Learning: A Prostate Cancer Case Study
Jacqueline J. Vallon, William Overman, Wanqiao Xu +11
Over the past decade, the use of machine learning (ML) models in healthcare applications has rapidly increased. Despite high performance, modern ML models do not always capture pat…
Scaling Clinician-Grade Feature Generation from Clinical Notes with Multi-Agent Language Models
Jiayi Wang, Jacqueline Jil Vallon, Nikhil V. Kotha +8
Developing accurate clinical prediction models is often bottlenecked by the difficulty of deriving meaningful structured features from unstructured EHR notes, a process that tradit…
Automated radiotherapy treatment planning guided by GPT-4Vision
Sheng Liu, Oscar Pastor-Serrano, Yizheng Chen +10
Objective: Radiotherapy treatment planning is a time-consuming and potentially subjective process that requires the iterative adjustment of model parameters to balance multiple con…