2 papers
cs.IR2026
Guaranteeing Faithful Evidence Extraction in Speculative Retrieval-Augmented Generation
Quentin Signé, Mohand Boughanem, Jose Moreno +1
Large Language Models (LLMs) are increasingly used as interfaces for information retrieval, but they remain prone to hallucinations and faithfulness errors, in which the generated…
cs.AI2020
Spot The Bot: A Robust and Efficient Framework for the Evaluation of Conversational Dialogue Systems
Jan Deriu, Don Tuggener, Pius von Däniken +6
The lack of time-efficient and reliable evaluation methods hamper the development of conversational dialogue systems (chatbots). Evaluations requiring humans to converse with chatb…