2 citations · 2 across the 1 of their papers we have counts for
1 paper
Dongseong Hwang, Weiran Wang, Zhuoyuan Huo +2
While Transformers have revolutionized deep learning, their quadratic attention complexity hinders their ability to process infinitely long inputs. We propose Feedback Attention Me…