8 citations · 10 across the 18 of their papers we have counts for
Showing cs.CRShow all
2 papers · 1 filter
cs.CR2026
DDSA: Dual-Domain Strategic Attack for Spatial-Temporal Efficiency in Adversarial Robustness Testing
Jinwei Hu, Shiyuan Meng, Yi Dong +1
Image transmission and processing systems in resource-critical applications face significant challenges from adversarial perturbations that compromise mission-specific object class…
cs.CR2024★ 8 cited
Safeguarding Large Language Models: A Survey
Yi Dong, Ronghui Mu, Yanghao Zhang +9
In the burgeoning field of Large Language Models (LLMs), developing a robust safety mechanism, colloquially known as "safeguards" or "guardrails", has become imperative to ensure t…