collaborators

5 papers

cs.SE2026

VulnGym: Benchmarking Coding Agents for Repository-Level Vulnerability Detection

Kexing Ji, Jiachen Liu, Enze Hu +7

Recent advances in LLM-based vulnerability detection have shown promising results, while coding agents further extend this capability from isolated code snippets to complete reposi…

cs.AI2026

Measuring Behavior Portability in Large Language Models

Tianjia Dong, Nadav Kunievsky, James A. Evans

Large language models are increasingly deployed as autonomous decision makers, yet the behavioral mapping they exhibit can vary substantially across decision environments that are…

cs.CR2026

DIPBox: A Multi-scale Testing Framework for Tracking Dataset Regeneration

Tian Dong, Yan Meng, Shaofeng Li +5

Training datasets have tremendous proprietary value and are vulnerable to unauthorized copying. Existing defenses mainly focus on tracking individual data points, but pay little at…

cs.CR2026

AgentRAE: Remote Action Execution through Notification-based Visual Backdoors against Screenshots-based Mobile GUI Agents

Yutao Luo, Haotian Zhu, Shuchao Pang +4

The rapid adoption of mobile graphical user interface (GUI) agents, which autonomously control applications and operating systems (OS), exposes new system-level attack surfaces. Ex…

cs.CR2025

Authority Backdoor: A Certifiable Backdoor Mechanism for Authoring DNNs

Han Yang, Shaofeng Li, Tian Dong +3

Deep Neural Networks (DNNs), as valuable intellectual property, face unauthorized use. Existing protections, such as digital watermarking, are largely passive; they provide only po…