2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.LG2024
RL in Markov Games with Independent Function Approximation: Improved Sample Complexity Bound under the Local Access Model
Junyi Fan, Yuxuan Han, Jialin Zeng +4
Efficiently learning equilibria with large state and action spaces in general-sum Markov games while overcoming the curse of multi-agency is a challenging problem. Recent works hav…
cs.LG2022★ 2 cited
Optimal Contextual Bandits with Knapsacks under Realizability via Regression Oracles
Yuxuan Han, Jialin Zeng, Yang Wang +2
We study the stochastic contextual bandit with knapsacks (CBwK) problem, where each action, taken upon a context, not only leads to a random reward but also costs a random resource…