Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
AgentAuditor: Human-Level Safety and Security Evaluation for LLM Agents
Hanjun Luo, Shenyu Dai, Chiming Ni +5
Despite the rapid advancement of LLM-based agents, the reliable evaluation of their safety and security remains a significant challenge. Existing rule-based or LLM-based evaluators…
cs.AI2025
Graph-Augmented Large Language Model Agents: Current Progress and Future Prospects
Yixin Liu, Guibin Zhang, Kun Wang +2
Autonomous agents based on large language models (LLMs) have demonstrated impressive capabilities in a wide range of applications, including web navigation, software development, a…