2 papers
cs.LG2025
Context Clues: Evaluating Long Context Models for Clinical Prediction Tasks on EHRs
Michael Wornow, Suhana Bedi, Miguel Angel Fuentes Hernandez +5
Foundation Models (FMs) trained on Electronic Health Records (EHRs) have achieved state-of-the-art results on numerous clinical prediction tasks. However, most existing EHR FMs hav…
cs.AI2024
WONDERBREAD: A Benchmark for Evaluating Multimodal Foundation Models on Business Process Management Tasks
Michael Wornow, Avanika Narayan, Ben Viggiano +15
Existing ML benchmarks lack the depth and diversity of annotations needed for evaluating models on business process management (BPM) tasks. BPM is the practice of documenting, meas…