1 paper
Moshe Kimhi, Tal Rozen, Avi Mendelson +1
Quantized neural networks are well known for reducing the latency, power consumption, and model size without significant harm to the performance. This makes them highly appropriate…