1 citations · 1 across the 3 of their papers we have counts for
1 paper · 1 filter
Doğaç Eldenk, Payal Mohapatra, Yigitcan Comlek +3
Speculative decoding accelerates LLM inference by drafting future tokens with a small model, but drafter models degrade sharply under template perturbation and long-context inputs.…