1 paper · 1 filter
Akul Malhotra, Sumeet Kumar Gupta
Ternary large language models (LLMs), which utilize ternary precision weights and 8-bit activations, have demonstrated competitive performance while significantly reducing the high…