12 papers
ClawSentry: A Progressive Multi-Tier Security Monitor for Safeguarding Autonomous LLM Agents
Kai Wang, Zeming Wei, BiaoJie Zeng +7
As large language model (LLM) agents move from conversation to executing code, reading local files, and orchestrating external tools, a single agent hijacked by a malicious third-p…
Deep Interaction: An Efficient Human-AI Interaction Method for Large Reasoning Models
Hefeng Zhou, Jinxuan Zhang, Jiong Lou +4
The paper introduces Deep Interaction, a method that lets users directly edit the chain‑of‑thought output of large language models to fix reasoning errors, resulting in higher corr…
Privacy-Preserving Text Sanitization for Distributed Agents Collaboration via Disentangled Representations
Xuan Liu, Hefeng Zhou, Sicheng Chen +6
When distributed agents exchange text across organizational boundaries, privacy leakage arises not only from explicit identifiers but also from distributional signatures such as fo…
AgentSchool: An LLM-Powered Multi-Agent Simulation for Education
Yulei Ye, Wenhao Li, Zhong Wen +23
Despite the rapid deployment of LLMs into classrooms, validating educational AI remains uniquely intractable: interventions act on developing learners whose cognitive and social tr…
SkillSafetyBench: Evaluating Agent Safety under Skill-Facing Attack Surfaces
Chang Jin, An Wang, Zeming Wei +7
Reusable skills are becoming a common interface for extending large language model agents, packaging procedural guidance with access to files, tools, memory, and execution environm…
TrinityGuard: A Unified Framework for Safeguarding Multi-Agent Systems
Kai Wang, Biaojie Zeng, Zeming Wei +7
With the rapid development of LLM-based multi-agent systems (MAS), their significant safety and security concerns have emerged, which introduce novel risks going beyond single agen…