1 paper
Rujikorn Charakorn, Edoardo Cetin, Shinnosuke Uesaka +1
Long input sequences are central to in-context learning, document understanding, and multi-step reasoning of Large Language Models (LLMs). However, the quadratic attention cost of…