3 citations · 3 across the 1 of their papers we have counts for
2 papers
cs.LG2024
GPT-2 Through the Lens of Vector Symbolic Architectures
Johannes Knittel, Tushaar Gangavarapu, Hendrik Strobelt +1
Understanding the general priniciples behind transformer models remains a complex endeavor. Experiments with probing and disentangling features using sparse autoencoders (SAE) sugg…
cs.CL2023★ 3 cited
The Role of Interactive Visualization in Explaining (Large) NLP Models: from Data to Inference
Richard Brath, Daniel Keim, Johannes Knittel +3
With a constant increase of learned parameters, modern neural language models become increasingly more powerful. Yet, explaining these complex model's behavior remains a widely uns…