1 citations · 1 across the 1 of their papers we have counts for
3 papers
cs.CR2025
Rethinking and Exploring String-Based Malware Family Classification in the Era of LLMs and RAG
Yufan Chen, Daoyuan Wu, Juantao Zhong +7
Malware family classification aims to identify the specific family (e.g., GuLoader or BitRAT) a malware sample may belong to, in contrast to malware detection or sample classificat…
cs.CR2024
SelfDefend: LLMs Can Defend Themselves against Jailbreaking in a Practical Manner
Xunguang Wang, Daoyuan Wu, Zhenlan Ji +7
Jailbreaking is an emerging adversarial attack that bypasses the safety alignment deployed in off-the-shelf large language models (LLMs) and has evolved into multiple categories: h…
cs.CR2024★ 1 cited
LLMs Can Defend Themselves Against Jailbreaking in a Practical Manner: A Vision Paper
Daoyuan Wu, Shuai Wang, Yang Liu +1
Jailbreaking is an emerging adversarial attack that bypasses the safety alignment deployed in off-the-shelf large language models (LLMs). A considerable amount of research exists p…