325 citations · 511 across the 28 of their papers we have counts for
11 papers · 2 filters
Neuron Interaction Based Representation Composition for Neural Machine Translation
Jian Li, Xing Wang, Baosong Yang +3
Recent NLP studies reveal that substantial linguistic information can be attributed to single neurons, i.e., individual dimensions of the representation vectors. We hypothesize tha…
Towards Understanding Neural Machine Translation with Word Importance
Shilin He, Zhaopeng Tu, Xing Wang +3
Although neural machine translation (NMT) has advanced the state-of-the-art on various language pairs, the interpretability of NMT remains unsatisfactory. In this work, we propose…
Multi-Granularity Self-Attention for Neural Machine Translation
Jie Hao, Xing Wang, Shuming Shi +2
Current state-of-the-art neural machine translation (NMT) uses a deep multi-head self-attention network with no explicit phrase information. However, prior work on statistical mach…
Towards Better Modeling Hierarchical Structure for Self-Attention with Ordered Neurons
Jie Hao, Xing Wang, Shuming Shi +2
Recent studies have shown that a hybrid of self-attention networks (SANs) and recurrent neural networks (RNNs) outperforms both individual architectures, while not much is known ab…
Self-Attention with Structural Position Representations
Xing Wang, Zhaopeng Tu, Longyue Wang +1
Although self-attention networks (SANs) have advanced the state-of-the-art on various NLP tasks, one criticism of SANs is their ability of encoding positions of input words (Shaw e…
One Model to Learn Both: Zero Pronoun Prediction and Translation
Longyue Wang, Zhaopeng Tu, Xing Wang +1
Zero pronouns (ZPs) are frequently omitted in pro-drop languages, but should be recalled in non-pro-drop languages. This discourse phenomenon poses a significant challenge for mach…