6 papers
Mathematical Principles and Experimental Discoveries of the Emergence of Symbolic Patterns in Artificial Neural Networks
Quanshi Zhang, Qihan Ren, Siyu Lou
Artificial Neural networks (ANNs) are often treated as black-box models, making explainability a central challenge in deep learning. Many engineering methods have been proposed to…
AgentDoG 1.5: A Lightweight and Scalable Alignment Framework for AI Agent Safety and Security
Dongrui Liu, Yu Li, Zhonghao Yang +47
Modern open-world agents such as OpenClaw exhibit powerful cross-environment execution capabilities yet introduce broad new safety risk sources. Meanwhile, advanced frontier AI mod…
The Interaction Bottleneck of Deep Neural Networks: Discovery, Proof, and Modulation
Huiqi Deng, Qihan Ren, Zhuofan Chen +5
Understanding what kinds of cooperative structures deep neural networks (DNNs) can represent remains a fundamental yet insufficiently understood problem. In this work, we treat int…
Evaluating the Correctness of Inference Patterns Used by LLMs for Judgment
Lu Chen, Yuxuan Huang, Yixing Li +6
This paper presents a method to analyze the inference patterns used by Large Language Models (LLMs) for judgment in a case study on legal LLMs, so as to identify potential incorrec…
Revisiting Generalization Power of a DNN in Terms of Symbolic Interactions
Lei Cheng, Junpeng Zhang, Qihan Ren +1
This paper aims to analyze the generalization power of deep neural networks (DNNs) from the perspective of interactions. Unlike previous analysis of a DNN's generalization power in…
Towards the Dynamics of a DNN Learning Symbolic Interactions
Qihan Ren, Junpeng Zhang, Yang Xu +3
This study proves the two-phase dynamics of a deep neural network (DNN) learning interactions. Despite the long disappointing view of the faithfulness of post-hoc explanation of a…