Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
Isolation as a First-Class Principle for LLM-Agent System Safety: Concepts, Taxonomy, Challenges and Future Directions
Huihao Jing, Wenbin Hu, Shaojin Chen +10
The capability of LLM agents to function as the ``brain'' of a system fundamentally expands the scope of analysis beyond a standalone model. Consequently, safety is no longer only…
cs.AI2026
SkillRevise: Improving LLM-Authored Agent Skills via Trace-Conditioned Skill Revision
Yuxuan Liu, Zhaochen Su, Lingyun Xie +11
Agent skills are procedural artifacts that enable LLM agents to execute workflows, verify constraints, and recover from failures. Existing self-evolving methods refine skills using…
cs.AI2025
MASLegalBench: Benchmarking Multi-Agent Systems in Deductive Legal Reasoning
Huihao Jing, Wenbin Hu, Hongyu Luo +4
Multi-agent systems (MAS), leveraging the remarkable capabilities of Large Language Models (LLMs), show great potential in addressing complex tasks. In this context, integrating MA…