Publications (77)
Co-PatcheR: Collaborative Software Patching with Component(s)-specific Small Reasoning Models
Yuheng Tang, Hongwei Li, Kaijie Zhu +3
QPanda: high-performance quantum computing framework for multiple application scenarios
Menghan Dou, Tianrui Zou, Yuan Fang +8
When Do Intrinsic Rewards Work for Code Reasoning? A Comprehensive Study
Xiaolong Jin, Xuandong Zhao, Wenbo Guo +2
Berkeley Open Extended Reality Recordings 2023 (BOXRR-23): 4.7 Million Motion Capture Recordings from 105,852 Extended Reality Device Users
Vivek Nair, Wenbo Guo, Rui Wang +3
Agents' Last Exam
Yiyou Sun, Xinyang Han, Weichen Zhang +306
netFound: Principled Design for Network Foundation Models
Sylee Beltiukov, Satyandra Guthula, Haarika Manda +5
F-Fidelity: A Robust Framework for Faithfulness Evaluation of Explainable AI
Xu Zheng, Farhad Shirani, Zhuomin Chen +4
RL-JACK: Reinforcement Learning-powered Black-box Jailbreaking Attack against LLMs
Xuan Chen, Yuzhou Nie, Lu Yan +3
TextGuard: Provable Defense against Backdoor Attacks on Text Classification
Hengzhi Pei, Jinyuan Jia, Wenbo Guo +2
ShieldNet: Network-Level Guardrails against Emerging Supply-Chain Injections in Agentic Systems
Zhuowen Yuan, Zhaorun Chen, Zhen Xiang +5
MAStrike: Shapley-Guided Collusive Red-Teaming on Multi-Agent Systems
Chejian Xu, Zhaorun Chen, Jingyang Zhang +5
rePIRL: Learn PRM with Inverse RL for LLM Reasoning
Xian Wu, Kaijie Zhu, Ying Zhang +2
BlockScan: Detecting Anomalies in Blockchain Transactions
Jiahao Yu, Xian Wu, Hao Liu +2
Demystifying Network Foundation Models
Sylee Beltiukov, Satyandra Guthula, Wenbo Guo +2
CyberGym-E2E: Scalable Real-World Benchmark for AI Agents' End-to-End Cybersecurity Capabilities
Tianneng Shi, Robin Rheem, Dongwei Jiang +13
Temporal Logic-Based Multi-Vehicle Backdoor Attacks against Offline RL Agents in End-to-end Autonomous Driving
Xuan Chen, Shiwei Feng, Zikang Xiong +6
MiqroForge: An Intelligent Workflow Platform for Quantum-Enhanced Computational Chemistry
Jianan Wang, Wenbo Guo, Xin Yue +6
HarnessLLM: Automatic Testing Harness Generation via Reinforcement Learning
Yujian Liu, Jiabao Ji, Yang Zhang +3
Learning Adversary-Resistant Deep Neural Networks
Qinglong Wang, Wenbo Guo, Kaixuan Zhang +4
In Search of netUnicorn: A Data-Collection Platform to Develop Generalizable ML Models for Network Security Problems
Roman Beltiukov, Wenbo Guo, Arpit Gupta +1
ExploitGym: Can AI Agents Turn Security Vulnerabilities into Real Attacks?
Zhun Wang, Nico Schiller, Hongwei Li +13
Cutting the Gordian Knot: Detecting Malicious PyPI Packages via a Knowledge-Mining Framework
Wenbo Guo, Chengwei Liu, Ming Kang +5
Inferring Private Personal Attributes of Virtual Reality Users from Head and Hand Motion Data
Vivek Nair, Christian Rack, Wenbo Guo +8
Explaining Deep Learning Models - A Bayesian Non-parametric Approach
Wenbo Guo, Sui Huang, Yunzhe Tao +2
To Defend Against Cyber Attacks, We Must Teach AI Agents to Hack
Terry Yue Zhuo, Yangruibo Ding, Wenbo Guo +1
SkillClone: Multi-Modal Clone Detection and Clone Propagation Analysis in the Agent Skill Ecosystem
Jiaying Zhu, Lyuye Zhang, Wenbo Guo +1
FORAY: Towards Effective Attack Synthesis against Deep Logical Vulnerabilities in DeFi Protocols
Hongbo Wen, Hanzhi Liu, Jiaxin Song +3
Mind the Inconspicuous: Revealing the Hidden Weakness in Aligned LLMs' Refusal Boundaries
Jiahao Yu, Haozheng Luo, Jerry Yao-Chieh Hu +3
High-Speed (7,2) Compressor Using A Fast Carry-Generation Logic based on Sorting Network
Wenbo Guo
Adversary Resistant Deep Neural Networks with an Application to Malware Detection
Qinglong Wang, Wenbo Guo, Kaixuan Zhang +4
Understanding NPM Malicious Package Detection: A Benchmark-Driven Empirical Analysis
Wenbo Guo, Zhongwen Chen, Zhengzi Xu +7
TABOR: A Highly Accurate Approach to Inspecting and Restoring Trojan Backdoors in AI Systems
Wenbo Guo, Lun Wang, Xinyu Xing +2
DANCE: Enhancing saliency maps using decoys
Yang Lu, Wenbo Guo, Xinyu Xing +1
DevOps-Gym: Benchmarking AI Agents in Software DevOps Cycle
Yuheng Tang, Kaijie Zhu, Bonan Ruan +14
Casting a SPELL: Sentence Pairing Exploration for LLM Limitation-breaking
Yifan Huang, Xiaojun Jia, Wenbo Guo +4
VulnLLM-R: Specialized Reasoning LLM with Agent Scaffold for Vulnerability Detection
Yuzhou Nie, Hongwei Li, Chengquan Guo +5
BandFuzz: An ML-powered Collaborative Fuzzing Framework
Wenxuan Shi, Hongwei Li, Jiahao Yu +3
Program Analysis Guided LLM Agent for Proof-of-Concept Generation
Achintya Desai, Md Shafiuzzaman, Wenbo Guo +1
The Attack and Defense Landscape of Agentic AI: A Comprehensive Survey
Juhee Kim, Xiaoyuan Liu, Zhun Wang +4
Data-driven analysis of the electronic-structure factors controlling the work functions of perovskite oxides
Yihuang Xiong, Weinan Chen, Wenbo Guo +2
MalwarePT: A Binary-Level Foundation Model for Malware Analysis
Saastha Vasan, Yuzhou Nie, Kaie Chen +6
An Empirical Study of Malicious Code In PyPI Ecosystem
Wenbo Guo, Zhengzi Xu, Chengwei Liu +3
Mutate to Bypass: Autonomous Endpoint Evasion via Knowledge-Driven Multi-Agent Orchestration
Weifeng Yuan, Wenbo Guo, Qingyun Du +4
AgentVigil: Generic Black-Box Red-teaming for Indirect Prompt Injection against LLM Agents
Zhun Wang, Vincent Siu, Zhe Ye +6
Bridging Expert Reasoning and LLM Detection: A Knowledge-Driven Framework for Malicious Packages
Wenbo Guo, Shiwen Song, Jiaxun Guo +5
BACKDOORL: Backdoor Attack against Competitive Reinforcement Learning
Lun Wang, Zaynah Javed, Xian Wu +3
SpearBot: Leveraging Large Language Models in a Generative-Critique Framework for Spear-Phishing Email Generation
Qinglin Qi, Yun Luo, Yijia Xu +2
Unique Identification of 50,000+ Virtual Reality Users from Head & Hand Motion Data
Vivek Nair, Wenbo Guo, Justus Mattern +4
PatchPilot: A Cost-Efficient Software Engineering Agent with Early Attempts on Formal Verification
Hongwei Li, Yuheng Tang, Shiqi Wang +1
Are Shortest Rationales the Best Explanations for Human Understanding?
Hua Shen, Tongshuang Wu, Wenbo Guo +1
LeakAgent: RL-based Red-teaming Agent for LLM Privacy Leakage
Yuzhou Nie, Zhun Wang, Ye Yu +4
OpenSage: Self-programming Agent Generation Engine
Hongwei Li, Zhun Wang, Qinrun Dai +11
DecodingTrust-Agent Platform (DTap): A Controllable and Interactive Red-Teaming Platform for AI Agents
Zhaorun Chen, Xun Liu, Haibo Tong +14
High-speed and high-efficiency three-dimensional shape measurement based on Gray-coded light
Zhoujie Wu, Wenbo Guo, Yueyang Li +2
Progent: Securing AI Agents with Privilege Control
Tianneng Shi, Jingxuan He, Zhun Wang +4
Seeing Is Not Screening: Multimodal Hidden Instruction Attacks on Agent Skill Scanners
Xiaojun Jia, Jie Liao, Simeng Qin +5
TermiGen: High-Fidelity Environment and Robust Trajectory Synthesis for Terminal Agents
Kaijie Zhu, Yuzhou Nie, Yijiang Li +10
Guiding Symbolic Execution with Static Analysis and LLMs for Vulnerability Discovery
Md Shafiuzzaman, Achintya Desai, Wenbo Guo +1
Using Non-invertible Data Transformations to Build Adversarial-Robust Neural Networks
Qinglong Wang, Wenbo Guo, Alexander G. Ororbia +6
IntelliRadar: A Comprehensive Platform to Pinpoint Malicious Package Information from Cyber Intelligence
Wenbo Guo, Chengwei Liu, Limin Wang +4
AgentBeats: Agentifying Agent Assessment for Openness, Standardization, and Reproducibility
Xiaoyuan Liu, Jianhong Tu, Yuqi Chen +26
SynthChain: A Synthetic Benchmark and Forensic Analysis of Advanced and Stealthy Software Supply Chain Attacks
Zhuoran Tan, Wenbo Guo, Taylor Brierley +3
PromptArmor: Simple yet Effective Prompt Injection Defenses
Tianneng Shi, Kaijie Zhu, Zhun Wang +13
3CAD: A Large-Scale Real-World 3C Product Dataset for Unsupervised Anomaly
Enquan Yang, Peng Xing, Hanyang Sun +4
When LLM Meets DRL: Advancing Jailbreaking Efficiency via DRL-guided Search
Xuan Chen, Yuzhou Nie, Wenbo Guo +1
Data Free Backdoor Attacks
Bochuan Cao, Jinyuan Jia, Chuxuan Hu +5
Deep Motion Masking for Secure, Usable, and Scalable Real-Time Anonymization of Virtual Reality Motion Data
Vivek Nair, Wenbo Guo, James F. O'Brien +2
Towards Interrogating Discriminative Machine Learning Models
Wenbo Guo, Kaixuan Zhang, Lin Lin +2
MalSkillBench: A Runtime-Verified Benchmark of Malicious Agent Skills
Wenbo Guo, Wei Zeng, Chengwei Liu +5
BlueCodeAgent: A Blue Teaming Agent Enabled by Automated Red Teaming for CodeGen AI
Chengquan Guo, Yuzhou Nie, Chulin Xie +3
From CVE Entries to Verifiable Exploits: An Automated Multi-Agent Framework for Reproducing CVEs
Saad Ullah, Praneeth Balasubramanian, Wenbo Guo +5
SeCodePLT: A Unified Platform for Evaluating the Security of Code GenAI
Yuzhou Nie, Zhun Wang, Yu Yang +7
Skills That Don't Exist: A Large-Scale Study of Hallucinated Skill Recommendation in LLM Agents
Weifeng Yuan, Wenbo Guo, Feng Dong +2
The paper investigates how large language model agents often fabricate nonexistent skill names when asked to recommend and install skills, exposing a supply‑chain security risk, an…
SeedSmith: LLM-Driven Seed Synthesis for Directed Fuzzing
Junmin Zhu, Siyu Liu, Jie Hu +7
The paper introduces SeedSmith, an LLM‑driven pipeline that automatically creates input seeds targeting specific sink functions and crash preconditions, improving the effectiveness…
Evolaris: A Roadmap to Self-Evolving Software Intelligence Management
Chengwei Liu, Wenbo Guo, Yuxin Zhang +4
Frontier AI's Impact on the Cybersecurity Landscape
Yujin Potter, Wenbo Guo, Zhun Wang +6
MELON: Provable Defense Against Indirect Prompt Injection Attacks in AI Agents
Kaijie Zhu, Xianjun Yang, Jindong Wang +2