activity
20152022
most citedDeploying a Steered Query Optimizer in Production at Microsoft

20 citations · 36 across the 10 of their papers we have counts for

collaborators

12 papers

cs.LG2022

Towards Data-Driven Offline Simulations for Online Reinforcement Learning

Shengpu Tang, Felipe Vieira Frujeri, Dipendra Misra +4

Modern decision-making systems, from robots to web recommendation engines, are expected to adapt: to user preferences, changing circumstances or even new tasks. Yet, it is still un…

cs.DB202220 cited

Deploying a Steered Query Optimizer in Production at Microsoft

Wangda Zhang, Matteo Interlandi, Paul Mineiro +6

Modern analytical workloads are highly heterogeneous and massively complex, making generic query optimizers untenable for many customers and scenarios. As a result, it is important…

stat.ML2022

A lower confidence sequence for the changing mean of non-negative right heavy-tailed observations with bounded mean

Paul Mineiro

A confidence sequence (CS) is an anytime-valid sequential inference primitive which produces an adapted sequence of sets for a predictable parameter sequence with a time-uniform co…

cs.LG20211 cited

Interaction-Grounded Learning

Tengyang Xie, John Langford, Paul Mineiro +1

Consider a prosthetic arm, learning to adapt to its user's control signals. We propose Interaction-Grounded Learning for this novel setting, in which a learner's goal is to interac…

cs.LG2021

ChaCha for Online AutoML

Qingyun Wu, Chi Wang, John Langford +2

We propose the ChaCha (Champion-Challengers) algorithm for making an online choice of hyperparameters in online learning settings. ChaCha handles the process of determining a champ…

cs.LG20212 cited

Improving Long-Term Metrics in Recommendation Systems using Short-Horizon Reinforcement Learning

Bogdan Mazoure, Paul Mineiro, Pavithra Srinath +3

We study session-based recommendation scenarios where we want to recommend items to users during sequential interactions to improve their long-term utility. Optimizing a long-term…