1.4k citations · 1.4k across the 9 of their papers we have counts for
10 papers
Provable Membership Inference Privacy
Zachary Izzo, Jinsung Yoon, Sercan O. Arik +1
In applications involving sensitive data, such as finance and healthcare, the necessity for preserving data privacy can be a significant barrier to machine learning model developme…
A Spectral Method for Assessing and Combining Multiple Data Visualizations
Rong Ma, Eric D. Sun, James Zou
Dimension reduction and data visualization aim to project a high-dimensional dataset to a low-dimensional space while capturing the intrinsic structures in the data. It is an indis…
SEAL : Interactive Tool for Systematic Error Analysis and Labeling
Nazneen Rajani, Weixin Liang, Lingjiao Chen +2
With the advent of Transformers, large language models (LLMs) have saturated well-known NLP benchmarks and leaderboards with high aggregate performance. However, many times these m…
C-Mixup: Improving Generalization in Regression
Huaxiu Yao, Yiping Wang, Linjun Zhang +2
Improving the generalization of deep networks is an important open challenge, particularly in domains without plentiful data. The mixup algorithm improves generalization by linearl…
Knowledge-Driven New Drug Recommendation
Zhenbang Wu, Huaxiu Yao, Zhe Su +5
Drug recommendation assists doctors in prescribing personalized medications to patients based on their health conditions. Existing drug recommendation solutions adopt the supervise…
Data Budgeting for Machine Learning
Xinyi Zhao, Weixin Liang, James Zou
Data is the fuel powering AI and creates tremendous value for many domains. However, collecting datasets for AI is a time-consuming, expensive, and complicated endeavor. For practi…