1 paper · 1 filter
Pál Zsámboki, Ádám Fraknói, Máté Gedeon +2
We present an in-depth mechanistic interpretability analysis of training small transformers on an elementary task, counting, which is a crucial deductive step in many algorithms. I…