From the 1 of 9 linked papers with an AI index.
9 papers
Auditing Emergent LLM-Agent Collaboration through Cooperation-Obligation Coupling
Zuyuan Zhang, Hanqing Yang, Carlee Joe-Wong +1
The paper proposes iCORE, a unified representation that combines a cooperation graph, an obligation graph, and an audit map to let auditors verify that each step of an LLM‑agent wo…
Operator-Guided Invariance Learning for Continuous Reinforcement Learning
Zuyuan Zhang, Fei Xu Yu, Tian Lan
Reinforcement learning (RL) with continuous time and state/action spaces is often data-intensive and brittle under nuisance variability and shift, motivating methods that exploit v…
LiSFC-Search: Lifelong Search for Network SFC Optimization under Non-stationary Drifts
Zuyuan Zhang, Vaneet Aggarwal, Tian Lan
Edge-cloud convergence is reshaping service provisioning across 5G/6G and computing power networks (CPNs). Service function chaining (SFC) requires continuously placing and schedul…
Cochain Perspectives on Temporal-Difference Signals for Learning Beyond Markov Dynamics
Zuyuan Zhang, Sizhe Tang, Tian Lan
Non-Markovian dynamics are commonly found in real-world environments due to long-range dependencies, partial observability, and memory effects. The Bellman equation that is the cen…
Structuring Value Representations via Geometric Coherence in Markov Decision Processes
Zuyuan Zhang, Zeyu Fang, Tian Lan
Geometric properties can be leveraged to stabilize and speed reinforcement learning. Existing examples include encoding symmetry structure, geometry-aware data augmentation, and en…
Manifold-Constrained Energy-Based Transition Models for Offline Reinforcement Learning
Zeyu Fang, Zuyuan Zhang, Mahdi Imani +1
Model-based offline reinforcement learning is brittle under distribution shift: policy improvement drives rollouts into state--action regions weakly supported by the dataset, where…