1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Zhaozhuo Xu, Minghao Yan, Junyan Zhang +1
Transformer models have demonstrated superior performance in natural language processing. The dot product self-attention in Transformer allows us to model interactions between word…