2 papers
cs.SE2026
Results-Actionability Gap: Understanding How Practitioners Evaluate LLM Products in the Wild
Willem van der Maden, Malak Sadek, Ziang Xiao +3
How do product teams evaluate LLM-powered products? As organizations integrate large language models (LLMs) into digital products, their unpredictable nature makes traditional eval…
cs.HC2024
Modulating Language Model Experiences through Frictions
Katherine M. Collins, Valerie Chen, Ilia Sucholutsky +6
Language models are transforming the ways that their users engage with the world. Despite impressive capabilities, over-consumption of language model outputs risks propagating unch…