Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
Joint Value Estimation and Bidding in Repeated First-Price Auctions
Yuxiao Wen, Yanjun Han, Zhengyuan Zhou
We study regret minimization in repeated first-price auctions (FPAs), where a bidder observes only the realized outcome after each auction -- win or loss. This setup reflects pract…
cs.LG2025
Optimal Arm Elimination Algorithms for Combinatorial Bandits
Yuxiao Wen, Yanjun Han, Zhengyuan Zhou
Combinatorial bandits extend the classical bandit framework to settings where the learner selects multiple arms in each round, motivated by applications such as online recommendati…
cs.LG2024
Stochastic contextual bandits with graph feedback: from independence number to MAS number
Yuxiao Wen, Yanjun Han, Zhengyuan Zhou
We consider contextual bandits with graph feedback, a class of interactive learning problems with richer structures than vanilla contextual bandits, where taking an action reveals…