2 papers
cs.CL2026
ConvApparel: A Benchmark Dataset and Validation Framework for User Simulators in Conversational Recommenders
Ofer Meshi, Krisztian Balog, Sally Goldman +5
The promise of LLM-based user simulators to improve conversational AI is hindered by a critical "realism gap," leading to systems that are optimized for simulated interactions, but…
cs.CL2025
"Stupid robot, I want to speak to a human!" User Frustration Detection in Task-Oriented Dialog Systems
Mireia Hernandez Caralt, Ivan SekuliÄ, Filip CareviÄ +9
Detecting user frustration in modern-day task-oriented dialog (TOD) systems is imperative for maintaining overall user satisfaction, engagement, and retention. However, most recent…