10 papers
FlowScout: From Execution Feedback to Reliable Tool-Using Agent Workflows
Shuo Hao, You Lu, Bihuan Chen +1
Agentic workflows have become an important abstraction for building reliable LLM-based automation systems by organizing large language models (LLMs), tools, and control logic into…
DOCSCHISEL: Adaptive Tool Documentation Optimization Framework for LLM Agents
You Lu, Kun Zhang, Bihuan Chen +1
Large language models (LLMs) increasingly rely on external tools to accomplish complex real-world tasks, making tool documentation a critical grounding resource for LLM agents. Exi…
SkillSentry: Reliable Skill Execution for LLM Agents via Runtime Assurance
You Lu, Xinyu Huang, Bihuan Chen +1
LLM agents are increasingly equipped with skills to perform complex tasks through multi-step reasoning and tool use. Although skills provide reusable procedural knowledge, agents m…
DynNPC: Finding More Violations Induced by ADS in Simulation Testing through Dynamic NPC Behavior Generation
You Lu, Yifan Tian, Dingji Wang +2
Recently, a number of simulation testing approaches have been proposed to generate diverse driving scenarios for autonomous driving systems (ADSs) testing. However, the behaviors o…
AutoMerge: Search-Based Model Merging Framework for Effective Model Reuse
You Lu, Jiyang Zhang, Bihuan Chen +3
Software reuse has long been recognized as a critical and widely studied topic in software engineering, offering substantial benefits in reducing development costs, improving softw…
A Survey of Fuzzing Open-Source Operating Systems
Kun Hu, Qicai Chen, Wenzhuo Zhang +7
Vulnerabilities in open-source operating systems (OSs) pose substantial security risks to software systems, making their detection crucial. While fuzzing has been an effective vuln…