Showing cs.LGShow all
2 papers · 1 filter
cs.LG2023
Compressed Real Numbers for AI: a case-study using a RISC-V CPU
Federico Rossi, Marco Cococcioni, Roger Ferrer Ibàñez +5
As recently demonstrated, Deep Neural Networks (DNN), usually trained using single precision IEEE 754 floating point numbers (binary32), can also work using lower precision. Theref…
cs.LG2018
Low-Precision Floating-Point Schemes for Neural Network Training
Marc Ortiz, Adrián Cristal, Eduard Ayguadé +1
The use of low-precision fixed-point arithmetic along with stochastic rounding has been proposed as a promising alternative to the commonly used 32-bit floating point arithmetic to…