23 citations · 23 across the 2 of their papers we have counts for
2 papers
cs.CL2023
Quantifying Context Mixing in Transformers
Hosein Mohebbi, Willem Zuidema, Grzegorz Chrupała +1
Self-attention weights and their transformed variants have been the main source of information for analyzing token-to-token interactions in Transformer-based models. But despite th…
cs.CL2016★ 23 cited
From phonemes to images: levels of representation in a recurrent neural model of visually-grounded language learning
Lieke Gelderloos, Grzegorz Chrupała
We present a model of visually-grounded language learning based on stacked gated recurrent neural networks which learns to predict visual features given an image description in the…