1 citations · 1 across the 1 of their papers we have counts for
3 papers
cs.GT2023★ 2 cited
A Finite-Sample Analysis of Payoff-Based Independent Learning in Zero-Sum Stochastic Games
Zaiwei Chen, Kaiqing Zhang, Eric Mazumdar +2
We study two-player zero-sum stochastic games, and propose a form of independent learning dynamics called Doubly Smoothed Best-Response dynamics, which integrates a discrete and do…
cs.LG2022
A Note on Zeroth-Order Optimization on the Simplex
Tijana Zrnic, Eric Mazumdar
We construct a zeroth-order gradient estimator for a smooth function defined on the probability simplex. The proposed estimator queries the simplex only. We prove that projected gr…
cs.LG2020★ 1 cited
Technical Report: Adaptive Control for Linearizable Systems Using On-Policy Reinforcement Learning
Tyler Westenbroek, Eric Mazumdar, David Fridovich-Keil +3
This paper proposes a framework for adaptively learning a feedback linearization-based tracking controller for an unknown system using discrete-time model-free policy-gradient para…