activity
20242026
collaborators
Showing cs.CRShow all

6 papers · 1 filter

cs.CR2026

ActBench: Self-Evolving Benchmark of Behavioral Safety in Cowork Agents

Hongwei Yao, Yiming Liu, Meihui Chen +6

Cowork agents may complete benign tasks while disclosing protected data, manipulating unauthorized state, invocate unauthorized API. We define behavioral safety and introduce ActBe…

cs.CR2025

ShadowCode: Towards (Automatic) External Prompt Injection Attack against Code LLMs

Yuchen Yang, Yiming Li, Hongwei Yao +5

Recent advancements have led to the widespread adoption of code-oriented large language models (Code LLMs) for programming tasks. Despite their success in deployment, their securit…

cs.CR2025

Explainer-guided Targeted Adversarial Attacks against Binary Code Similarity Detection Models

Mingjie Chen, Tiancheng Zhu, Mingxue Zhang +4

Binary code similarity detection (BCSD) serves as a fundamental technique for various software engineering tasks, e.g., vulnerability detection and classification. Attacks against…

cs.CR2025

Combating Concept Drift with Explanatory Detection and Adaptation for Android Malware Classification

Yiling He, Junchi Lei, Zhan Qin +2

Machine learning-based Android malware classifiers achieve high accuracy in stationary environments but struggle with concept drift. The rapid evolution of malware, especially with…

cs.CR2024

DeUEDroid: Detecting Underground Economy Apps Based on UTG Similarity

Zhuo Chen, Jie Liu, Yubo Hu +7

In recent years, the underground economy is proliferating in the mobile system. These underground economy apps (UEware) make profits from providing non-compliant services, especial…

cs.CR2024

Explanation as a Watermark: Towards Harmless and Multi-bit Model Ownership Verification via Watermarking Feature Attribution

Shuo Shao, Yiming Li, Hongwei Yao +3

Ownership verification is currently the most critical and widely adopted post-hoc method to safeguard model copyright. In general, model owners exploit it to identify whether a giv…