1 citations · 1 across the 3 of their papers we have counts for
3 papers
PairSAE: Mechanistic Interpretability from Pair Representations in Protein Co-Folding
Giosue Migliorini, Aristofanis Rontogiannis, Grigori Guitchounts +3
Foundation models for structural biology have achieved remarkable performance in predicting biomolecular structure and show promise for the design of proteins and small molecules.…
Dynamics of the Transformer Residual Stream: Coupling Spectral Geometry to Network Topology
Jesseba Fernando, Grigori Guitchounts
Large language models are remarkably capable, yet how computation propagates through their layers remains poorly understood. A growing line of work treats depth as discrete time an…
Transformer Dynamics: A neuroscientific approach to interpretability of large language models
Jesseba Fernando, Grigori Guitchounts
As artificial intelligence models have exploded in scale and capability, understanding of their internal mechanisms remains a critical challenge. Inspired by the success of dynamic…