From the 2 of 16 linked papers with an AI index.
16 papers
FlowScout: From Execution Feedback to Reliable Tool-Using Agent Workflows
Shuo Hao, You Lu, Bihuan Chen +1
Agentic workflows have become an important abstraction for building reliable LLM-based automation systems by organizing large language models (LLMs), tools, and control logic into…
DOCSCHISEL: Adaptive Tool Documentation Optimization Framework for LLM Agents
You Lu, Kun Zhang, Bihuan Chen +1
Large language models (LLMs) increasingly rely on external tools to accomplish complex real-world tasks, making tool documentation a critical grounding resource for LLM agents. Exi…
SkillSentry: Reliable Skill Execution for LLM Agents via Runtime Assurance
You Lu, Xinyu Huang, Bihuan Chen +1
LLM agents are increasingly equipped with skills to perform complex tasks through multi-step reasoning and tool use. Although skills provide reusable procedural knowledge, agents m…
ProfMalPlus: Agent-Coordinated Detection of Malicious NPM Packages via Static-Dynamic Analysis Synergy
Yiheng Huang, Zhijia Zhao, Bihuan Chen +6
The paper presents ProfMalPlus, a system that detects malicious NPM packages by combining object-sensitive behavior graphs with coordinated large language model reasoning over stat…
VulWeaver: Weaving Broken Semantics for Grounded Vulnerability Detection
Yiheng Cao, Yihao Chen, Xin Hu +9
VulWeaver is an LLM‑driven system that improves source‑code vulnerability detection by combining deterministic static analysis with LLM‑based semantic inference to build a unified…
DynNPC: Finding More Violations Induced by ADS in Simulation Testing through Dynamic NPC Behavior Generation
You Lu, Yifan Tian, Dingji Wang +2
Recently, a number of simulation testing approaches have been proposed to generate diverse driving scenarios for autonomous driving systems (ADSs) testing. However, the behaviors o…