1 paper
Andrey Gromov, Kushal Tirumala, Hassan Shapourian +2
How is knowledge stored in an LLM's weights? We study this via layer pruning: if removing a certain layer does not affect model performance in common question-answering benchmarks,…