2 citations · 2 across the 3 of their papers we have counts for
3 papers
cs.LG2026
Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs
Meichen Song, Yuhao Wang, Enlu Zhou
In online reinforcement learning, data scarcity creates epistemic uncertainty that makes robustness important early in learning, whereas sufficient exploration is needed to learn t…
math.OC2023★ 2 cited
Quantile Optimization via Multiple Timescale Local Search for Black-box Functions
Jiaqiao Hu, Meichen Song, Michael C. Fu
We consider quantile optimization of black-box functions that are estimated with noise. We propose two new iterative three-timescale local search algorithms. The first algorithm us…
cs.MA2023
Differentiable Arbitrating in Zero-sum Markov Games
Jing Wang, Meichen Song, Feng Gao +3
We initiate the study of how to perturb the reward in a zero-sum Markov game with two players to induce a desirable Nash equilibrium, namely arbitrating. Such a problem admits a bi…