1 paper · 1 filter
David Bani-Harouni, Chantal Pellegrini, Paul Stangel +4
A safe and trustworthy use of Large Language Models (LLMs) requires an accurate expression of confidence in their answers. We propose a novel Reinforcement Learning approach that a…