1 citations · 1 across the 1 of their papers we have counts for
2 papers
cs.LG2022★ 1 cited
Learning Representation from Neural Fisher Kernel with Low-rank Approximation
Ruixiang Zhang, Shuangfei Zhai, Etai Littwin +1
In this paper, we study the representation of neural networks from the view of kernels. We first define the Neural Fisher Kernel (NFK), which is the Fisher Kernel applied to neural…
cs.LG2021
An Attention Free Transformer
Shuangfei Zhai, Walter Talbott, Nitish Srivastava +4
We introduce Attention Free Transformer (AFT), an efficient variant of Transformers that eliminates the need for dot product self attention. In an AFT layer, the key and value are…