Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
Psychological Influences of Conversational AI: Research and Design Directions for Reducing Harm and Promoting Well-Being
Jina Suh, Mihaela Vorvoreanu, Forough Poursabzi-Sangdeh +5
As conversational AI systems become increasingly integrated into daily life, their potential effects on user well-being require ongoing attention. While consumer-facing generalist…
cs.AI2026
DisaBench: A Participatory Evaluation Framework for Disability Harms in Language Models
Eugenia Kim, Ioana Tanase, Christina Mallon
General-purpose safety benchmarks for large language models do not adequately evaluate disability-related harms. We introduce DisaBench: a taxonomy of twelve disability harm catego…
cs.AI2025
Lessons From Red Teaming 100 Generative AI Products
Blake Bullwinkel, Amanda Minnich, Shiven Chawla +23
In recent years, AI red teaming has emerged as a practice for probing the safety and security of generative AI systems. Due to the nascency of the field, there are many open questi…