849 citations · 2.2k across the 17 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2023★ 22 cited
Debiasing Vision-Language Models via Biased Prompts
Ching-Yao Chuang, Varun Jampani, Yuanzhen Li +2
Machine learning models have been shown to inherit biases from their training datasets. This can be particularly problematic for vision-language foundation models trained on uncura…
cs.LG2021★ 4 cited
Editing a classifier by rewriting its prediction rules
Shibani Santurkar, Dimitris Tsipras, Mahalaxmi Elango +3
We present a methodology for modifying the behavior of a classifier by directly rewriting its prediction rules. Our approach requires virtually no additional data collection and ca…