1 citations · 1 across the 1 of their papers we have counts for
1 paper
Oscar Brown, Zhengjie Wang, Andrea Do +2
The acceleration of Large Language Models (LLMs) with speculative decoding provides a significant runtime improvement without any loss of accuracy. Currently, EAGLE-2 is the state-…