2 citations · 2 across the 1 of their papers we have counts for
2 papers
cs.LG2021★ 2 cited
Translational Equivariance in Kernelizable Attention
Max Horn, Kumar Shridhar, Elrich Groenewald +1
While Transformer architectures have show remarkable success, they are bound to the computation of all pairwise interactions of input element and thus suffer from limited scalabili…
cs.LG2019
ProbAct: A Probabilistic Activation Function for Deep Neural Networks
Kumar Shridhar, Joonho Lee, Hideaki Hayashi +6
Activation functions play an important role in training artificial neural networks. The majority of currently used activation functions are deterministic in nature, with their fixe…