1 paper · 1 filter
Ting-Yun Chang, Jesse Thomason, Robin Jia
This paper studies in-context learning by decomposing the output of large language models into the individual contributions of attention heads and MLPs (components). We observe cur…