7 citations · 7 across the 3 of their papers we have counts for
1 paper · 1 filter
Jialiang Kang, Han Shu, Wenshuo Li +2
Speculative decoding is a widely adopted technique for accelerating inference in large language models (LLMs), yet its application to vision-language models (VLMs) remains underexp…