22 citations · 47 across the 9 of their papers we have counts for
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2024★ 2 cited
Using LLMs to discover emerging coded antisemitic hate-speech in extremist social media
Dhanush Kikkisetti, Raza Ul Mustafa, Wendy Melillo +4
Online hate speech proliferation has created a difficult problem for social media platforms. A particular challenge relates to the use of coded language by groups interested in bot…
cs.CL2023
Robust Infidelity: When Faithfulness Measures on Masked Language Models Are Misleading
Evan Crothers, Herna Viktor, Nathalie Japkowicz
A common approach to quantifying neural text classifier interpretability is to calculate faithfulness metrics based on iteratively masking salient input tokens and measuring change…