3 papers
cs.LG2026
Decentralized Best-Response-Based Learning in Two-Player Zero-Sum Stochastic Games: A Finite-Sample Analysis
Zaiwei Chen, Kaiqing Zhang, Eric Mazumdar +2
We present a finite-sample analysis of decentralized learning in two-player zero-sum matrix games and stochastic games, with a focus on best-response-based learning algorithms. In…
cs.LG2025
Reinforcement Learning with Imperfect Transition Predictions: A Bellman-Jensen Approach
Chenbei Lu, Zaiwei Chen, Tongxin Li +2
Traditional reinforcement learning (RL) assumes the agents make decisions based on Markov decision processes (MDPs) with one-step transition models. In many real-world applications…
math.OC2025
Maximizing the Value of Predictions in Control: Accuracy Is Not Enough
Yiheng Lin, Christopher Yeh, Zaiwei Chen +1
We study the value of stochastic predictions in online optimal control with random disturbances. Prior work provides performance guarantees based on prediction error but ignores th…