2 papers
cs.LG2026
Tokenization Tradeoffs in Structured EHR Foundation Models
Lin Lawrence Guo, Santiago Eduardo Arciniegas, Joseph Jihyung Lee +4
Foundation models for structured electronic health records (EHRs) are pretrained on longitudinal sequences of timestamped clinical events to learn adaptable patient representations…
stat.ML2025
Understanding challenges to the interpretation of disaggregated evaluations of algorithmic fairness
Stephen R. Pfohl, Natalie Harris, Chirag Nagpal +12
Disaggregated evaluation across subgroups is critical for assessing the fairness of machine learning models, but its uncritical use can mislead practitioners. We show that equal pe…