Showing cs.CYShow all
2 papers · 1 filter
cs.CY2026
Beyond Simulations: What 20,000 Real Conversations Reveal About Mental Health AI Safety
Caitlin A. Stamatis, Jonah Meyerhoff, Richard Zhang +3
Mental-health AI safety is typically evaluated with small, simulation-based benchmarks that may not reflect the linguistic and contextual diversity of deployment. We pair four benc…
cs.CY2025
When Testing AI Tests Us: Safeguarding Mental Health on the Digital Frontlines
Sachin R. Pendse, Darren Gergle, Rachel Kornfield +6
Red-teaming is a core part of the infrastructure that ensures that AI models do not produce harmful content. Unlike past technologies, the black box nature of generative AI systems…