8 citations · 8 across the 1 of their papers we have counts for
3 papers
cs.LG2026
IoUCert: Robustness Verification for Anchor-based Object Detectors
Benedikt Brückner, Alejandro J. Mercado, Yanghao Zhang +2
While formal robustness verification has seen significant success in image classification, scaling these guarantees to object detection remains notoriously difficult due to complex…
cs.CR2025
Never compromise with vulnerabilities: a comprehensive survey on AI governance
Yuchu Jiang, Jian Zhao, Yuchen Yuan +62
The rapid advancement of AI has expanded its capabilities across domains, yet introduced critical technical vulnerabilities, such as algorithmic bias and adversarial sensitivity, t…
cs.CR2024★ 8 cited
Safeguarding Large Language Models: A Survey
Yi Dong, Ronghui Mu, Yanghao Zhang +9
In the burgeoning field of Large Language Models (LLMs), developing a robust safety mechanism, colloquially known as "safeguards" or "guardrails", has become imperative to ensure t…