2 citations · 2 across the 1 of their papers we have counts for
1 paper
Clemens JS Schaefer, Navid Lambert-Shirzad, Xiaofan Zhang +7
Efficiently serving neural network models with low latency is becoming more challenging due to increasing model complexity and parameter count. Model quantization offers a solution…