2 papers
cs.IR2026
Seeking Information with RAG-Assistants: Does Model Size Matter in Human-AI Collaborations?
Lennard C. Froma, Tom Kouwenhoven, Maaike H. T. de Boer +2
Much research on LLMs has focused on increasing benchmark performance. However, the evaluation of such models in real-world collaborative human-AI workflows has stayed behind. This…
cs.IR2026
"I Don't Know" -- Towards Appropriate Trust with Certainty-Aware Retrieval Augmented Generation
Daan Di Scala, Maaike de Boer, Pınar Yolum
Achieving the right amount of trust in AI systems is important, but challenging. The problem is exacerbated with the rise of Large Language Models (LLMs) as they provide human-leve…