2 papers
cs.AI2026
Expert Evaluation and the Limits of Human Feedback in Mental Health AI Safety Testing
Kiana Jafari, Paul Ulrich Nikolaus Rust, Duncan Eddy +7
Learning from human feedback~(LHF) assumes that expert judgments, appropriately aggregated, yield valid ground truth for training and evaluating AI systems. We tested this assumpti…
cs.HC2025
Seeking Late Night Life Lines: Experiences of Conversational AI Use in Mental Health Crisis
Leah Hope Ajmani, Arka Ghosh, Benjamin Kaveladze +7
Online, people often recount their experiences turning to conversational AI agents (e.g., ChatGPT, Claude, Copilot) for mental health support -- going so far as to replace their th…