4 papers · 1 filter
ProbGuard: Proactive Runtime Monitoring for LLM Agent Safety via Probabilistic Prediction
Haoyu Wang, Christopher M. Poskitt, Jiali Wei +1
Large Language Model (LLM) agents increasingly operate across domains such as robotics, virtual assistants, and web automation. However, their stochastic decision-making introduces…
Domain-Specialized Tree of Thought through Plug-and-Play Predictors
Xuanqi Gao, Haoyu Wang, Jun Sun +2
While Large Language Models (LLMs) have advanced complex reasoning, prominent methods like the Tree of Thoughts (ToT) framework face a critical trade-off between exploration depth…
Robust and Efficient Tool Orchestration via Layered Execution Structures with Reflective Correction
Tao Zhe, Haoyu Wang, Bo Luo +6
Tool invocation is a core capability of agentic systems, yet failures often arise not from individual tool calls but from how multiple tools are organized and executed together. Ex…
AgentSpec: Customizable Runtime Enforcement for Safe and Reliable LLM Agents
Haoyu Wang, Christopher M. Poskitt, Jun Sun
Agents built on LLMs are increasingly deployed across diverse domains, automating complex decision-making and task execution. However, their autonomy introduces safety risks, inclu…