2 citations · 2 across the 6 of their papers we have counts for
1 paper · 1 filter
Moran Yanuka, Paul Dixon, Eyal Finkelshtein +2
Speculative decoding accelerates autoregressive speech generation by letting a fast draft model propose tokens that a larger target model verifies. However, for speech LLMs that ge…