2 citations · 2 across the 3 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025★ 2 cited
Pruning Strategies for Backdoor Defense in LLMs
Santosh Chapagain, Shah Muhammad Hamdi, Soukaina Filali Boubrahimi
Backdoor attacks are a significant threat to the performance and integrity of pre-trained language models. Although such models are routinely fine-tuned for downstream NLP tasks, r…
cs.LG2025
Advancing Hate Speech Detection with Transformers: Insights from the MetaHate
Santosh Chapagain, Shah Muhammad Hamdi, Soukaina Filali Boubrahimi
Hate speech is a widespread and harmful form of online discourse, encompassing slurs and defamatory posts that can have serious social, psychological, and sometimes physical impact…