1 paper · 1 filter
Preetum Nakkiran, Arwen Bradley, Adam GoliÅski +3
Large Language Models (LLMs) often lack meaningful confidence estimates for their outputs. While base LLMs are known to exhibit next-token calibration, it remains unclear whether t…