1 paper · 1 filter
Dongseong Hwang, Weiran Wang, Zhuoyuan Huo +2
While Transformers have revolutionized deep learning, their quadratic attention complexity hinders their ability to process infinitely long inputs. We propose Feedback Attention Me…