activity
20242026
most citedEvaluating and Mitigating Linguistic Discrimination in Large Language Models

1 citations · 1 across the 9 of their papers we have counts for

collaborators
Showing cs.AIShow all

5 papers · 1 filter

cs.AI2026

MemGuard: Persisting Verifier Signals for LLM-Agent Memory Governance

Haoyu Wang, Guangyuan Dong, He Liang +5

LLM agents are moving from single-prompt use to long task streams in which reusable memory becomes a core capability for terminal, software-engineering, and web tasks. Such memory…

cs.AI2026

Domain-Specialized Tree of Thought through Plug-and-Play Predictors

Xuanqi Gao, Haoyu Wang, Jun Sun +2

While Large Language Models (LLMs) have advanced complex reasoning, prominent methods like the Tree of Thoughts (ToT) framework face a critical trade-off between exploration depth…

cs.AI2026

Robust and Efficient Tool Orchestration via Layered Execution Structures with Reflective Correction

Tao Zhe, Haoyu Wang, Bo Luo +6

Tool invocation is a core capability of agentic systems, yet failures often arise not from individual tool calls but from how multiple tools are organized and executed together. Ex…

cs.AI2025

ProbGuard: Proactive Runtime Monitoring for LLM Agent Safety via Probabilistic Prediction

Haoyu Wang, Christopher M. Poskitt, Jiali Wei +1

Large Language Model (LLM) agents increasingly operate across domains such as robotics, virtual assistants, and web automation. However, their stochastic decision-making introduces…

cs.AI2025

AgentSpec: Customizable Runtime Enforcement for Safe and Reliable LLM Agents

Haoyu Wang, Christopher M. Poskitt, Jun Sun

Agents built on LLMs are increasingly deployed across diverse domains, automating complex decision-making and task execution. However, their autonomy introduces safety risks, inclu…