collaborators

5 papers

cs.LG2026

Domain Knowledge Based Temporal-Spatial Graph Convolution Network for ECG Recognition

Wenting Ma, Zhipeng Zhang, Xiaohang Yuan +6

In light of strides in Arti cial Intelligence (AI) and its wide spread application, challenges persist in the interpretability of AI models, particularly within specialized domains…

cs.LG2026

Learning Can Converge Stably to the Wrong Belief under Latent Reliability

Zhipeng Zhang, Zhenjie Yao, Kai Li +1

Learning systems are typically optimized by minimizing loss or maximizing reward, assuming that improvements in these signals reflect progress toward the true objective. However, w…

cs.LG2026

Meta-Cognitive Reinforcement Learning with Self-Doubt and Recovery

Zhipeng Zhang, Xiongfei Su, Kai Li

Robust reinforcement learning methods typically focus on suppressing unreliable experiences or corrupted rewards, but they lack the ability to reason about the reliability of their…

cs.LG2026

Learning to Trust Experience: A Monitor-Trust-Regulator Framework for Learning under Unobservable Feedback Reliability

Zhipeng Zhang, Zhenjie Yao, Kai Li +1

Learning under unobservable feedback reliability poses a distinct challenge beyond optimization robustness: a system must decide whether to learn from an experience, not only how t…

cs.LG2026

Training instability in deep learning follows low-dimensional dynamical principles

Zhipeng Zhang, Zhenjie Yao, Kai Li +1

Deep learning systems achieve remarkable empirical performance, yet the stability of the training process itself remains poorly understood. Training unfolds as a high-dimensional d…