2 papers
cs.CL2025
Faithfulness metric fusion: Improving the evaluation of LLM trustworthiness across domains
Ben Malin, Tatiana Kalganova, Nikolaos Boulgouris
We present a methodology for improving the accuracy of faithfulness evaluation in Large Language Models (LLMs). The proposed methodology is based on the combination of elementary f…
cs.CL2024
A review of faithfulness metrics for hallucination assessment in Large Language Models
Ben Malin, Tatiana Kalganova, Nikoloas Boulgouris
This review examines the means with which faithfulness has been evaluated across open-ended summarization, question-answering and machine translation tasks. We find that the use of…