34 citations · 52 across the 3 of their papers we have counts for
3 papers
cs.LG2022★ 34 cited
ML Interpretability: Simple Isn't Easy
Tim Räz
The interpretability of ML models is important, but it is not clear what it amounts to. So far, most philosophers have discussed the lack of interpretability of black-box models su…
cs.LG2022★ 13 cited
Emergence of Concepts in DNNs?
Tim Räz
The present paper reviews and discusses work from computer science that proposes to identify concepts in internal representations (hidden layers) of DNNs. It is examined, first, ho…
cs.CY2022★ 5 cited
Gerrymandering Individual Fairness
Tim Räz
Individual fairness, proposed by Dwork et al., is a fairness measure that is supposed to prevent the unfair treatment of individuals on the subgroup level, and to overcome the prob…