papers

Publications (13)

cs.SE2025

Leveraging LLM Agents for Automated Video Game Testing

Chengjia Wang, Lanling Tang, Ming Yuan +3

cs.SE2026

Weaponizing the Commons: A Taxonomy and Detection Framework of Abuse on GitHub

Yuli Cheng, Xiaoyu Zhang, Jiongchi Yu +3

cs.CR2026

ARGUS: Defending LLM Agents Against Context-Aware Prompt Injection

Shihao Weng, Yang Feng, Jinrui Zhang +3

cs.SE2025

AutoEmpirical: LLM-Based Automated Research for Empirical Software Fault Analysis

Jiongchi Yu, Weipeng Jiang, Xiaoyu Zhang +3

cs.CR2025

CAShift: Benchmarking Log-Based Cloud Attack Detection under Normality Shift

Jiongchi Yu, Xiaofei Xie, Qiang Hu +6

cs.SE2026

Human in the Loop for Fuzz Testing: Literature Review and the Road Ahead

Jiongchi Yu, Xiaolin Wen, Sizhe Cheng +3

cs.AI2026

DynaTrust: Defending Multi-Agent Systems Against Sleeper Agents via Dynamic Trust Graphs

Yu Li, Qiang Hu, Yao Zhang +3

cs.CL2026

PTCBENCH: Benchmarking Contextual Stability of Personality Traits in LLM Systems

Jiongchi Yu, Yuhan Ma, Xiaoyu Zhang +4

cs.SE2025

Understanding the Supply Chain and Risks of Large Language Model Applications

Yujie Ma, Lili Quan, Xiaofei Xie +4

cs.SE2025

Defects4C: Benchmarking Large Language Model Repair Capability with C/C++ Bugs

Jian Wang, Xiaofei Xie, Qiang Hu +4

cs.SE2025

The Foundation Cracks: A Comprehensive Study on Bugs and Testing Practices in LLM Libraries

Weipeng Jiang, Xiaoyu Zhang, Xiaofei Xie +4

cs.SE2025

A.S.E: A Repository-Level Benchmark for Evaluating Security in AI-Generated Code

Keke Lian, Bin Wang, Lei Zhang +19

cs.CR2026

Chimera: Harnessing Multi-Agent LLMs for Automatic Insider Threat Simulation

Jiongchi Yu, Xiaofei Xie, Qiang Hu +2