1 paper · 1 filter
Zitong Li, Aparna Chandramowlishwaran
Sparse attention is a core building block in many leading neural network models, from graph-structured learning to sparse sequence modeling. It can be decomposed into a sequence of…