2 citations · 4 across the 9 of their papers we have counts for
1 paper · 1 filter
Dongseong Hwang, Weiran Wang, Zhuoyuan Huo +2
While Transformers have revolutionized deep learning, their quadratic attention complexity hinders their ability to process infinitely long inputs. We propose Feedback Attention Me…