2 papers
cs.HC2026
Is Seeing Believing? Evaluating Human Sensitivity to Synthetic Video
David Wegmann, Emil Stevnsborg, Søren Knudsen +2
Advances in machine learning have enabled the creation of realistic synthetic videos known as deepfakes. As deepfakes proliferate, concerns about rapid spread of disinformation and…
cs.SE2026
Results-Actionability Gap: Understanding How Practitioners Evaluate LLM Products in the Wild
Willem van der Maden, Malak Sadek, Ziang Xiao +3
How do product teams evaluate LLM-powered products? As organizations integrate large language models (LLMs) into digital products, their unpredictable nature makes traditional eval…