1 paper · 1 filter
Xiang Guan, Roger D. Newman-Norlund, Yong Yang +8
Mechanistic interpretability of large language models lacks spatially resolved, falsifiable tools for testing whether internal components are specialized for distinct cognitive ope…