1 paper · 1 filter
Jacob Nielsen, Peter Schneider-Kamp
Recently proposed methods for 1-bit and 1.58-bit quantization aware training investigate the performance and behavior of these methods in the context of large language models, find…