2 papers
cs.LG2026
Learning to Sparsify Stochastic Linear Bandits
Zhengmiao Wang, Ming Chi, Zhi-Wei Liu +2
This paper addresses the problem of learning to sparsify stochastic linear bandits, where a decision-maker sequentially selects actions from a high-dimensional space subject to a s…
math.OC2025
Online Convex Optimization with Memory and Limited Predictions
Zhengmiao Wang, Zhi-Wei Liu, Ming Chi +3
This paper addresses an online convex optimization problem where the cost function at each step depends on a history of past decisions (i.e., memory), and the decision maker has ac…