From the 1 of 1 linked paper with an AI index.
1 paper
Yuesong Liu, Yuan Zeng, Min Lyu +3
The paper proposes SparseSpec-L, a training-free self-speculative decoding method that uses a sparsified key‑value cache and an entropy‑based controller to speed up long‑context in…