Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Structure Before Collapse: Transient semantic geometry in next-token prediction
Yize Zhao, Isabel Papadimitriou, Christos Thrampoulidis
Neural Collapse predicts that balanced one-hot classification pushes model representations to be equally far from each other; a symmetric configuration that depends only on the out…
cs.LG2024
Using Shapley interactions to understand how models use structure
Divyansh Singhvi, Diganta Misra, Andrej Erkelens +3
Language is an intricately structured system, and a key goal of NLP interpretability is to provide methodological insights for understanding how language models represent this stru…