24 citations · 73 across the 4 of their papers we have counts for
1 paper · 1 filter
Nan Yang, Tao Ge, Liang Wang +5
We propose LLMA, an LLM accelerator to losslessly speed up Large Language Model (LLM) inference with references. LLMA is motivated by the observation that there are abundant identi…