3 papers
stat.ML2025
Sparsity-Based Interpolation of External, Internal and Swap Regret
Zhou Lu, Y. Jennifer Sun, Zhiyu Zhang
Focusing on the expert problem in online learning, this paper studies the interpolation of several performance metrics via -regret minimization, which measures the total loss o…
cs.LG2024
Second Order Methods for Bandit Optimization and Control
Arun Suggala, Y. Jennifer Sun, Praneeth Netrapalli +1
Bandit convex optimization (BCO) is a general framework for online decision making under uncertainty. While tight regret bounds for general convex losses have been established, exi…
cs.LG2024
Tight Rates for Bandit Control Beyond Quadratics
Y. Jennifer Sun, Zhou Lu
Unlike classical control theory, such as Linear Quadratic Control (LQC), real-world control problems are highly complex. These problems often involve adversarial perturbations, ban…