1 citations · 1 across the 3 of their papers we have counts for
5 papers
Oracle-Efficient Combinatorial Semi-Bandits
Jung-hun Kim, Milan Vojnović, Min-hwan Oh
We study the combinatorial semi-bandit problem where an agent selects a subset of base arms and receives individual feedback. While this generalizes the classical multi-armed bandi…
GL-LowPopArt: A Nearly Instance-Wise Minimax-Optimal Estimator for Generalized Low-Rank Trace Regression
Junghyun Lee, Kyoungseok Jang, Kwang-Sung Jun +2
We present `GL-LowPopArt`, a novel Catoni-style estimator for generalized low-rank trace regression. Building on `LowPopArt` (Jang et al., 2024), it employs a two-stage approach: n…
Combinatorial Bandits for Maximum Value Reward Function under Max Value-Index Feedback
Yiliu Wang, Wei Chen, Milan Vojnović
We consider a combinatorial multi-armed bandit problem for maximum value reward function under maximum value and index feedback. This is a new feedback structure that lies in betwe…
Doubly Adversarial Federated Bandits
Jialin Yi, Milan Vojnović
We study a new non-stochastic federated multi-armed bandit problem with multiple agents collaborating via a communication network. The losses of the arms are assigned by an oblivio…
Sketching stochastic valuation functions
Milan Vojnović, Yiliu Wang
We consider the problem of sketching set valuation functions, defined as the expectation of a valuation function applied to independent random item values. For valuation functions…