5 papers
SafeClaw-R: Towards Safe and Secure Multi-Agent Personal Assistants
Haoyu Wang, Zibo Xiao, Yedi Zhang +2
LLM-based multi-agent systems (MASs) are transforming personal productivity by autonomously executing complex, cross-platform tasks. Frameworks such as OpenClaw demonstrate the pot…
Domain-Specialized Tree of Thought through Plug-and-Play Predictors
Xuanqi Gao, Haoyu Wang, Jun Sun +2
While Large Language Models (LLMs) have advanced complex reasoning, prominent methods like the Tree of Thoughts (ToT) framework face a critical trade-off between exploration depth…
Robust and Efficient Tool Orchestration via Layered Execution Structures with Reflective Correction
Tao Zhe, Haoyu Wang, Bo Luo +6
Tool invocation is a core capability of agentic systems, yet failures often arise not from individual tool calls but from how multiple tools are organized and executed together. Ex…
LLM-enabled Applications Require System-Level Threat Monitoring
Yedi Zhang, Haoyu Wang, Xianglin Yang +2
LLM-enabled applications are rapidly reshaping the software ecosystem by using large language models as core reasoning components for complex task execution. This paradigm shift, h…
AgentSpec: Customizable Runtime Enforcement for Safe and Reliable LLM Agents
Haoyu Wang, Christopher M. Poskitt, Jun Sun
Agents built on LLMs are increasingly deployed across diverse domains, automating complex decision-making and task execution. However, their autonomy introduces safety risks, inclu…