1 paper
Jacob Nielsen, Peter Schneider-Kamp
Recently proposed methods for 1-bit and 1.58-bit quantization aware training investigate the performance and behavior of these methods in the context of large language models, find…