2 papers
cs.AI2026
Echoes of Human Malice in Agents: Benchmarking LLMs for Multi-Turn Online Harassment Attacks
Trilok Padhi, Pinxian Lu, Abdulkadir Erol +5
Large Language Model (LLM) agents are powering a growing share of interactive web applications, yet remain vulnerable to misuse and harm. Prior jailbreak research has largely focus…
cs.SI2026
Context-Aware Detection and Victim-Centered Response Generation for Online Harassment in Private Messaging
Pinxian Lu, Nimra Ishfaq, Emma Win +5
Online harassment is a widespread social and public health concern, yet most computational approaches for detecting and addressing harassment focus on publicly visible social media…