2 citations · 2 across the 1 of their papers we have counts for
1 paper
Alexander Modell, Patrick Rubin-Delanchy, Nick Whiteley
There is a large ongoing scientific effort in mechanistic interpretability to map embeddings and internal representations of AI systems into human-understandable concepts. A key el…