4 papers
Imperfectly Cooperative Human-AI Interactions: Comparing the Impacts of Human and AI Attributes in Simulated and User Studies
Myke C. Cohen, Mingqian Zheng, Neel Bhandari +6
AI design characteristics and human personality traits each impact the quality and outcomes of human-AI interactions. However, their relative and joint impacts are underexplored in…
Density-Guided Response Optimization: Community-Grounded Alignment via Implicit Acceptance Signals
Patrick Gerard, Svitlana Volkova
Language models deployed in online communities must adapt to norms that vary across social, cultural, and domain-specific contexts. Prior alignment approaches rely on explicit pref…
Proactive Defense: Compound AI for Detecting Persuasion Attacks and Measuring Inoculation Effectiveness
Svitlana Volkova, Will Dupree, Hsien-Te Kao +4
This paper introduces BRIES, a novel compound AI architecture designed to detect and measure the effectiveness of persuasion attacks across information environments. We present a s…
Building Resilient Information Ecosystems: Large LLM-Generated Dataset of Persuasion Attacks
Hsien-Te Kao, Aleksey Panasyuk, Peter Bautista +5
Organization's communication is essential for public trust, but the rise of generative AI models has introduced significant challenges by generating persuasive content that can for…