3 papers
cs.LG2025
Adaptive Client Sampling in Federated Learning via Online Learning with Bandit Feedback
Boxin Zhao, Lingxiao Wang, Ziqi Liu +4
Due to the high cost of communication, federated learning (FL) systems need to sample a subset of clients that are involved in each round of training. As a result, client sampling…
stat.ML2024
Instrumental Variable Value Iteration for Causal Offline Reinforcement Learning
Luofeng Liao, Zuyue Fu, Zhuoran Yang +3
In offline reinforcement learning (RL) an optimal policy is learned solely from a priori collected observational data. However, in observational data, actions are often confounded…
stat.ME2024
Confidence Sets for Causal Orderings
Y. Samuel Wang, Mladen Kolar, Mathias Drton
Causal discovery procedures aim to deduce causal relationships among variables in a multivariate dataset. While various methods have been proposed for estimating a single causal mo…