1 citations · 3 across the 10 of their papers we have counts for
1 paper · 1 filter
Yulong Huang, Xiang Liu, Hongxiang Huang +5
Linear Attention (LA) offers a promising paradigm for scaling large language models (LLMs) to long sequences by avoiding the quadratic complexity of self-attention. Recent LA model…