2 citations · 2 across the 3 of their papers we have counts for
3 papers
cs.CR2025
Towards Understanding the Safety Boundaries of DeepSeek Models: Evaluation and Findings
Zonghao Ying, Guangyi Zheng, Yongxin Huang +6
This study presents the first comprehensive safety evaluation of the DeepSeek models, focusing on evaluating the safety risks associated with their generated content. Our evaluatio…
cs.CL2024★ 2 cited
Multi-Turn Context Jailbreak Attack on Large Language Models From First Principles
Xiongtao Sun, Deyue Zhang, Dongdong Yang +2
Large language models (LLMs) have significantly enhanced the performance of numerous applications, from intelligent conversations to text generation. However, their inherent securi…
cs.SE2023
ConFL: Constraint-guided Fuzzing for Machine Learning Framework
Zhao Liu, Quanchen Zou, Tian Yu +4
As machine learning gains prominence in various sectors of society for automated decision-making, concerns have risen regarding potential vulnerabilities in machine learning (ML) f…