8 citations · 8 across the 2 of their papers we have counts for
2 papers
cs.CL2025
Breaking Down Bias: On The Limits of Generalizable Pruning Strategies
Sibo Ma, Alejandro Salinas, Peter Henderson +1
We employ model pruning to examine how LLMs conceptualize racial biases, and whether a generalizable mitigation strategy for such biases appears feasible. Our analysis yields sever…
cs.CL2024★ 8 cited
What's in a Name? Auditing Large Language Models for Race and Gender Bias
Alejandro Salinas, Amit Haim, Julian Nyarko
We employ an audit design to investigate biases in state-of-the-art large language models, including GPT-4. In our study, we prompt the models for advice involving a named individu…