1 paper · 1 filter
Bokun Wang, Axel Berg, Durmus Alp Emre Acar +1
Recent work has shown that 8-bit floating point (FP8) can be used for efficiently training neural networks with reduced computational cost compared to training in FP32/FP16. In thi…