Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
Cracking the Code: Enhancing Implicit Hate Speech Detection through Coding Classification
Lu Wei, Liangzhi Li, Tong Xiang +2
The internet has become a hotspot for hate speech (HS), threatening societal harmony and individual well-being. While automatic detection methods perform well in identifying explic…
cs.CL2024
Imposter.AI: Adversarial Attacks with Hidden Intentions towards Aligned Large Language Models
Xiao Liu, Liangzhi Li, Tong Xiang +4
With the development of large language models (LLMs) like ChatGPT, both their vast applications and potential vulnerabilities have come to the forefront. While developers have inte…