1 paper
Yongwan Jo, Jinyoung Park, Euihyun Lee +1
Modern large language models (LLMs) exhibit activation sparsity, wherein only a subset of their neurons is activated for given input tokens. Researchers have leveraged this propert…