Publications (37)
Boundary-targeted Membership Inference Attacks on Safety Classifiers
Anthony Hughes, Alexander Goldberg, Prince Jha +3
Safety classifiers are essential safeguards within generative AI systems, filtering harmful content or identifying at-risk users when interacting with large language models. Despit…
Vipera: Blending Visual and LLM-Driven Guidance for Systematic Auditing of Text-to-Image Generative AI
Yanwei Huang, Wesley Hanwen Deng, Sijia Xiao +4
Despite their increasing capabilities, text-to-image generative AI systems are known to produce biased, offensive, and otherwise problematic outputs. While recent advancements have…
An Interactive Interpretability System for Breast Cancer Screening with Deep Learning
Yuzhe Lu, Adam Perer
Deep learning methods, in particular convolutional neural networks, have emerged as a powerful tool in medical image computing tasks. While these complex models provide excellent p…
Extended Analysis of "How Child Welfare Workers Reduce Racial Disparities in Algorithmic Decisions"
Logan Stapleton, Hao-Fei Cheng, Anna Kawakami +7
This is an extended analysis of our paper "How Child Welfare Workers Reduce Racial Disparities in Algorithmic Decisions," which looks at racial disparities in the Allegheny Family…
StructVizor: Interactive Profiling of Semi-Structured Textual Data
Yanwei Huang, Yan Miao, Di Weng +2
Data profiling plays a critical role in understanding the structure of complex datasets and supporting numerous downstream tasks, such as social media analytics and financial fraud…
Texture: Structured Exploration of Text Datasets
Will Epperson, Arpit Mathur, Adam Perer +1
Exploratory analysis of a text corpus is essential for assessing data quality and developing meaningful hypotheses. Text analysis relies on understanding documents through structur…