62 citations · 151 across the 15 of their papers we have counts for
27 papers
Probabilistic Missing Value Imputation for Mixed Categorical and Ordered Data
Yuxuan Zhao, Alex Townsend, Madeleine Udell
Many real-world datasets contain missing entries and mixed data types including categorical and ordered (e.g. continuous and ordinal) variables. Imputing the missing entries is nec…
gcimpute: A Package for Missing Data Imputation
Yuxuan Zhao, Madeleine Udell
This article introduces the Python package gcimpute for missing data imputation. gcimpute can impute missing data with many different variable types, including continuous, binary,…
Towards Group Robustness in the presence of Partial Group Labels
Vishnu Suresh Lokhande, Kihyuk Sohn, Jinsung Yoon +3
Learning invariant representations is an important requirement when training machine learning models that are driven by spurious correlations in the datasets. These spurious correl…
ControlBurn: Feature Selection by Sparse Forests
Brian Liu, Miaolan Xie, Madeleine Udell
Tree ensembles distribute feature importance evenly amongst groups of correlated features. The average feature ranking of the correlated group is suppressed, which reduces interpre…
Privileged Zero-Shot AutoML
Nikhil Singh, Brandon Kates, Jeff Mentch +3
This work improves the quality of automated machine learning (AutoML) systems by using dataset and function descriptions while significantly decreasing computation time from minute…
Tensor Random Projection for Low Memory Dimension Reduction
Yiming Sun, Yang Guo, Joel A. Tropp +1
Random projections reduce the dimension of a set of vectors while preserving structural information, such as distances between vectors in the set. This paper proposes a novel use o…