Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Dynamics Models for Offline Hyperparameter Selection in Real-World RL
Jordan Coblin, Han Wang, Martha White +1
A key obstacle to deploying reinforcement learning in real-world systems is hyperparameter selection, particularly when simulators are unavailable and online experimentation is cos…
cs.LG2024
Federated Temporal Difference Learning with Linear Function Approximation under Environmental Heterogeneity
Han Wang, Aritra Mitra, Hamed Hassani +2
We initiate the study of federated reinforcement learning under environmental heterogeneity by considering a policy evaluation problem. Our setup involves agents interacting wi…