2 citations · 4 across the 3 of their papers we have counts for
3 papers
Supporting Human Raters with the Detection of Harmful Content using Large Language Models
Kurt Thomas, Patrick Gage Kelley, David Tao +7
In this paper, we explore the feasibility of leveraging large language models (LLMs) to automate or otherwise assist human raters with identifying harmful content including hate sp…
Understanding Help-Seeking and Help-Giving on Social Media for Image-Based Sexual Abuse
Miranda Wei, Sunny Consolvo, Patrick Gage Kelley +7
Image-based sexual abuse (IBSA), like other forms of technology-facilitated abuse, is a growing threat to people's digital safety. Attacks include unwanted solicitations for sexual…
Robust, privacy-preserving, transparent, and auditable on-device blocklisting
Kurt Thomas, Sarah Meiklejohn, Michael A. Specter +4
With the accelerated adoption of end-to-end encryption, there is an opportunity to re-architect security and anti-abuse primitives in a manner that preserves new privacy expectatio…