1 paper
Martina G. Vilas, Federico Adolfi, David Poeppel +1
Inner Interpretability is a promising emerging field tasked with uncovering the inner mechanisms of AI systems, though how to develop these mechanistic theories is still much debat…