4 papers
ActBench: Self-Evolving Benchmark of Behavioral Safety in Cowork Agents
Hongwei Yao, Yiming Liu, Meihui Chen +6
Cowork agents may complete benign tasks while disclosing protected data, manipulating unauthorized state, invocate unauthorized API. We define behavioral safety and introduce ActBe…
ShadowCode: Towards (Automatic) External Prompt Injection Attack against Code LLMs
Yuchen Yang, Yiming Li, Hongwei Yao +5
Recent advancements have led to the widespread adoption of code-oriented large language models (Code LLMs) for programming tasks. Despite their success in deployment, their securit…
Explainer-guided Targeted Adversarial Attacks against Binary Code Similarity Detection Models
Mingjie Chen, Tiancheng Zhu, Mingxue Zhang +4
Binary code similarity detection (BCSD) serves as a fundamental technique for various software engineering tasks, e.g., vulnerability detection and classification. Attacks against…
Combating Concept Drift with Explanatory Detection and Adaptation for Android Malware Classification
Yiling He, Junchi Lei, Zhan Qin +2
Machine learning-based Android malware classifiers achieve high accuracy in stationary environments but struggle with concept drift. The rapid evolution of malware, especially with…