works on

From the 2 of 17 linked papers with an AI index.

activity
20242026
collaborators

17 papers

cs.LG2026

FlowScout: From Execution Feedback to Reliable Tool-Using Agent Workflows

Shuo Hao, You Lu, Bihuan Chen +1

Agentic workflows have become an important abstraction for building reliable LLM-based automation systems by organizing large language models (LLMs), tools, and control logic into…

cs.LG2026

DOCSCHISEL: Adaptive Tool Documentation Optimization Framework for LLM Agents

You Lu, Kun Zhang, Bihuan Chen +1

Large language models (LLMs) increasingly rely on external tools to accomplish complex real-world tasks, making tool documentation a critical grounding resource for LLM agents. Exi…

cs.AI2026

SkillSentry: Reliable Skill Execution for LLM Agents via Runtime Assurance

You Lu, Xinyu Huang, Bihuan Chen +1

LLM agents are increasingly equipped with skills to perform complex tasks through multi-step reasoning and tool use. Although skills provide reusable procedural knowledge, agents m…

cs.SE2026

ProfMalPlus: Agent-Coordinated Detection of Malicious NPM Packages via Static-Dynamic Analysis Synergy

Yiheng Huang, Zhijia Zhao, Bihuan Chen +6

The paper presents ProfMalPlus, a system that detects malicious NPM packages by combining object-sensitive behavior graphs with coordinated large language model reasoning over stat…

cs.SE2026

VulWeaver: Weaving Broken Semantics for Grounded Vulnerability Detection

Yiheng Cao, Yihao Chen, Xin Hu +9

VulWeaver is an LLM‑driven system that improves source‑code vulnerability detection by combining deterministic static analysis with LLM‑based semantic inference to build a unified…

cs.SE2026

DynNPC: Finding More Violations Induced by ADS in Simulation Testing through Dynamic NPC Behavior Generation

You Lu, Yifan Tian, Dingji Wang +2

Recently, a number of simulation testing approaches have been proposed to generate diverse driving scenarios for autonomous driving systems (ADSs) testing. However, the behaviors o…