Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
AgentDoG 1.5: A Lightweight and Scalable Alignment Framework for AI Agent Safety and Security
Dongrui Liu, Yu Li, Zhonghao Yang +47
Modern open-world agents such as OpenClaw exhibit powerful cross-environment execution capabilities yet introduce broad new safety risk sources. Meanwhile, advanced frontier AI mod…
cs.AI2025
Evaluating the Correctness of Inference Patterns Used by LLMs for Judgment
Lu Chen, Yuxuan Huang, Yixing Li +6
This paper presents a method to analyze the inference patterns used by Large Language Models (LLMs) for judgment in a case study on legal LLMs, so as to identify potential incorrec…