4 papers
SNIP: An Adaptation of Sorted Neighborhood Methods for Deduplicating Pedigree Data
Theodore Huang, Matthew Ploenzke, Danielle Braun
Pedigree data contain family history information that is used to analyze hereditary diseases. These clinical data sets may contain duplicate records due to the same family visiting…
Extending Models Via Gradient Boosting: An Application to Mendelian Models
Theodore Huang, Gregory Idos, Christine Hong +3
Improving existing widely-adopted prediction models is often a more efficient and robust way towards progress than training new models from scratch. Existing models may (a) incorpo…
PanelPRO: A R package for multi-syndrome, multi-gene risk modeling for individuals with a family history of cancer
Gavin Lee, Qing Zhang, Jane W. Liang +4
Identifying individuals who are at high risk of cancer due to inherited germline mutations is critical for effective implementation of personalized prevention strategies. Most exis…
Combining Breast Cancer Risk Prediction Models
Zoe Guan, Theodore Huang, Anne Marie McCarthy +6
Accurate risk stratification is key to reducing cancer morbidity through targeted screening and preventative interventions. Numerous breast cancer risk prediction models have been…