3 papers
cs.LG2026
Decision-Weighted Flow Matching for Contextual Stochastic Optimization
Jize Xie, Haomiao Wu, Qiang Chen +2
Conditional generative models are increasingly used as scenario generators for stochastic optimization, but standard training objectives emphasize uniform distributional fit rather…
cs.LG2025
Cascading Bandits Robust to Adversarial Corruptions
Jize Xie, Cheng Chen, Zhiyong Wang +1
Online learning to rank sequentially recommends a small list of items to users from a large candidate set and receives the users' click feedback. In many real-world scenarios, user…
cs.LG2025
Online Clustering of Dueling Bandits
Zhiyong Wang, Jiahang Sun, Mingze Kong +4
The contextual multi-armed bandit (MAB) is a widely used framework for problems requiring sequential decision-making under uncertainty, such as recommendation systems. In applicati…