3 citations · 3 across the 1 of their papers we have counts for
2 papers
cs.LG2026
Measuring Model Robustness via Fisher Information: Spectral Bounds, Theoretical Guarantees, and Practical Algorithms
Chong Zhang, Xiang Li, Jia Wang +2
The robustness of deep neural networks is crucial for safety-critical deployments, yet existing evaluation methods are often attack-dependent and lack interpretability. We propose…
cs.CL2024★ 3 cited
Target-driven Attack for Large Language Models
Chong Zhang, Mingyu Jin, Dong Shu +3
Current large language models (LLM) provide a strong foundation for large-scale user-oriented natural language tasks. Many users can easily inject adversarial text or instructions…