From the 1 of 3 linked papers with an AI index.
3 papers
cs.LG2026
Minimal Markovization via Stable Quotients in Holonomy-Cover Decision Processes
Zuyuan Zhang, Yongshan Chen, Mahdi Imani +1
The paper defines the smallest memory representation needed for a class of partially observable decision processes called holonomy-cover decision processes, builds a stable quotien…
cs.LG2026
FedQHD: Closed-Form Function-Space Federated Reinforcement Learning
Yuchen Hou, Yongshan Chen, Zhuowen Zou +4
Federated reinforcement learning enables decentralized agents to collaboratively improve policies or value estimates without exchanging raw trajectories. However, FedAvg-style para…
cs.LG2026
Online Learning and Equilibrium Computation with Ranking Feedback
Mingyang Liu, Yongshan Chen, Zhiyuan Fan +3
Online learning in arbitrary, and possibly adversarial, environments has been extensively studied in sequential decision-making, and it is closely connected to equilibrium computat…