3 papers
cs.HC2026
When Chatbots Accommodate: What AI Companions Optimize for in Vulnerable Conversations
Minh Duc Chu, Yifan Wu, Zhiyi Chen +2
Millions turn to AI companion chatbots during loneliness, grief, and personal crises. How these companion platforms respond in such moments can shape the trajectory of a user's vul…
cs.CL2026
How Far Will They Go? Red-Teaming Online Influence with Large Language Models
Daniel C. Ruiz, Anna Serbina, Ashwin Rao +2
As large language model (LLM)-based agents increasingly participate in online discourse, red-teaming their capacity to support political influence campaigns is critical for informa…
cs.LG2026
Graph-Regularized Sparse Autoencoders for LLM Safety Steering
Jehyeok Yeon, Federico Cinus, Yifan Wu +1
Sparse autoencoders (SAEs) are increasingly used to extract activation directions for inference-time steering, but their standard sparsity objective treats latent features as indep…