38 citations · 50 across the 10 of their papers we have counts for
10 papers
CrowdCounter: A benchmark type-specific multi-target counterspeech dataset
Punyajoy Saha, Abhilash Datta, Abhik Jana +1
Counterspeech presents a viable alternative to banning or suspending users for hate speech while upholding freedom of expression. However, writing effective counterspeech is challe…
Demarked: A Strategy for Enhanced Abusive Speech Moderation through Counterspeech, Detoxification, and Message Management
Seid Muhie Yimam, Daryna Dementieva, Tim Fischer +8
Despite regulations imposed by nations and social media platforms, such as recent EU regulations targeting digital violence, abusive content persists as a significant challenge. Ex…
On Zero-Shot Counterspeech Generation by LLMs
Punyajoy Saha, Aalok Agrawal, Abhik Jana +2
With the emergence of numerous Large Language Models (LLM), the usage of such models in various Natural Language Processing (NLP) applications is increasing extensively. Counterspe…
InfFeed: Influence Functions as a Feedback to Improve the Performance of Subjective Tasks
Somnath Banerjee, Maulindu Sarkar, Punyajoy Saha +2
Recently, influence functions present an apparatus for achieving explainability for deep neural models by quantifying the perturbation of individual train instances that might impa…
Low-Resource Counterspeech Generation for Indic Languages: The Case of Bengali and Hindi
Mithun Das, Saurabh Kumar Pandey, Shivansh Sethi +2
With the rise of online abuse, the NLP community has begun investigating the use of neural architectures to generate counterspeech that can "counter" the vicious tone of such abusi…
Probing LLMs for hate speech detection: strengths and vulnerabilities
Sarthak Roy, Ashish Harshavardhan, Animesh Mukherjee +1
Recently efforts have been made by social media platforms as well as researchers to detect hateful or toxic language using large language models. However, none of these works aim t…