Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Learning to Select and Rank from Choice-Based Feedback: A Simple Nested Approach
Junwen Yang, Yifan Feng
We study a ranking and selection problem of learning from choice-based feedback with dynamic assortments. In this problem, a company sequentially displays a set of items to a popul…
cs.LG2024
Optimal Clustering with Bandit Feedback
Junwen Yang, Zixin Zhong, Vincent Y. F. Tan
This paper considers the problem of online clustering with bandit feedback. A set of arms (or items) can be partitioned into various groups that are unknown. Within each group, the…