1 paper
Karen Jia-Hui Li, Simone Balloccu, Ondrej Dusek +1
The increasing trust in large language models (LLMs), especially in the form of chatbots, is often undermined by the lack of their extrinsic evaluation. This holds particularly tru…