1 paper
Pedram Zamirai, Jian Zhang, Christopher R. Aberger +1
State-of-the-art generic low-precision training algorithms use a mix of 16-bit and 32-bit precision, creating the folklore that 16-bit hardware compute units alone are not enough t…