distributionally robust reinforcement learning 1exploration-exploitation tradeoff 1interactive data collection 1robust inventory control 1robust markov decision processes 1sample complexity 1
From the 1 of 5 linked papers with an AI index.
Showing stat.MLShow all
2 papers · 1 filter
stat.ML2026
Robust Assortment Optimization from Observational Data
Miao Lu, Yuxuan Han, Han Zhong +2
Assortment optimization is a fundamental challenge in modern retail and recommendation systems, where the goal is to select a subset of products that maximizes expected revenue und…
stat.ML2025
Learning an Optimal Assortment Policy under Observational Data
Yuxuan Han, Han Zhong, Miao Lu +2
We study the fundamental problem of offline assortment optimization under the Multinomial Logit (MNL) model, where sellers must determine the optimal subset of the products to offe…