Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Agentao: A Policy-Governed Runtime Harness for Embeddable Tool-Using LLM Agents
Bo Jin, Qiang Jiao, Xin Tong
LLM agents increasingly operate as execution systems that invoke tools, modify local state, use persistent memory, and interact with external protocols. These capabilities make age…
cs.AI2024
CPSDBench: A Large Language Model Evaluation Benchmark and Baseline for Chinese Public Security Domain
Xin Tong, Bo Jin, Zhi Lin +3
Large Language Models (LLMs) have demonstrated significant potential and effectiveness across multiple application domains. To assess the performance of mainstream LLMs in public s…