1 paper
Anton Rasmussen, Hong Qin
Quantized large language models enable on-premises processing of sensitive data, but their confidence estimates must be trustworthy. Reliability depends on implementation choices--…