2 papers
cs.LG2026
Functional multi-armed bandit and the best function identification problems
Yuriy Dorn, Aleksandr Katrutsa, Ilgam Latypov +1
Bandit optimization usually refers to the class of online optimization problems with limited feedback, namely, a decision maker uses only the objective value at the current point t…
cs.LG2026
UCB-type Algorithm for Budget-Constrained Expert Learning
Ilgam Latypov, Alexandra Suvorikova, Alexey Kroshnin +2
In many modern applications, a system must dynamically choose between several adaptive learning algorithms that are trained online. Examples include model selection in streaming en…