Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Minimal Markovization via Stable Quotients in Holonomy-Cover Decision Processes
Zuyuan Zhang, Yongshan Chen, Mahdi Imani +1
An agent acting under partial observability must retain a recursively updateable statistic of history that restores the Markov property, but the smallest such statistic is generall…
cs.LG2026
FedQHD: Closed-Form Function-Space Federated Reinforcement Learning
Yuchen Hou, Yongshan Chen, Zhuowen Zou +4
Federated reinforcement learning enables decentralized agents to collaboratively improve policies or value estimates without exchanging raw trajectories. However, FedAvg-style para…