5 citations · 10 across the 5 of their papers we have counts for
1 paper · 2 filters
Xiaofan Lu, Yixiao Zeng, Feiyang Ma +2
Speculative Decoding (SD) is a technique to accelerate the inference of Large Language Models (LLMs) by using a lower complexity draft model to propose candidate tokens verified by…