9 citations · 18 across the 2 of their papers we have counts for
3 papers · 1 filter
Understanding Multi-Head Attention in Abstractive Summarization
Joris Baan, Maartje ter Hoeve, Marlies van der Wees +2
Attention mechanisms in deep learning architectures have often been used as a means of transparency and, as such, to shed light on the inner workings of the architectures. Recently…
Do Transformer Attention Heads Provide Transparency in Abstractive Summarization?
Joris Baan, Maartje ter Hoeve, Marlies van der Wees +2
Learning algorithms become more powerful, often at the cost of increased complexity. In response, the demand for algorithms to be transparent is growing. In NLP tasks, attention di…
Dynamic Data Selection for Neural Machine Translation
Marlies van der Wees, Arianna Bisazza, Christof Monz
Intelligent selection of training data has proven a successful technique to simultaneously increase training efficiency and translation performance for phrase-based machine transla…