papers
Publications (13)
cs.SE2025
Leveraging LLM Agents for Automated Video Game Testing
Chengjia Wang, Lanling Tang, Ming Yuan +3
cs.SE2026
Weaponizing the Commons: A Taxonomy and Detection Framework of Abuse on GitHub
Yuli Cheng, Xiaoyu Zhang, Jiongchi Yu +3
cs.CR2026
ARGUS: Defending LLM Agents Against Context-Aware Prompt Injection
Shihao Weng, Yang Feng, Jinrui Zhang +3
cs.SE2025
AutoEmpirical: LLM-Based Automated Research for Empirical Software Fault Analysis
Jiongchi Yu, Weipeng Jiang, Xiaoyu Zhang +3
cs.CR2025
CAShift: Benchmarking Log-Based Cloud Attack Detection under Normality Shift
Jiongchi Yu, Xiaofei Xie, Qiang Hu +6
cs.SE2026
Human in the Loop for Fuzz Testing: Literature Review and the Road Ahead
Jiongchi Yu, Xiaolin Wen, Sizhe Cheng +3
cs.AI2026
DynaTrust: Defending Multi-Agent Systems Against Sleeper Agents via Dynamic Trust Graphs
Yu Li, Qiang Hu, Yao Zhang +3
cs.CL2026
PTCBENCH: Benchmarking Contextual Stability of Personality Traits in LLM Systems
Jiongchi Yu, Yuhan Ma, Xiaoyu Zhang +4
cs.SE2025
Understanding the Supply Chain and Risks of Large Language Model Applications
Yujie Ma, Lili Quan, Xiaofei Xie +4
cs.SE2025
Defects4C: Benchmarking Large Language Model Repair Capability with C/C++ Bugs
Jian Wang, Xiaofei Xie, Qiang Hu +4
cs.SE2025
The Foundation Cracks: A Comprehensive Study on Bugs and Testing Practices in LLM Libraries
Weipeng Jiang, Xiaoyu Zhang, Xiaofei Xie +4
cs.SE2025
A.S.E: A Repository-Level Benchmark for Evaluating Security in AI-Generated Code
Keke Lian, Bin Wang, Lei Zhang +19
cs.CR2026
Chimera: Harnessing Multi-Agent LLMs for Automatic Insider Threat Simulation
Jiongchi Yu, Xiaofei Xie, Qiang Hu +2