2 citations · 2 across the 1 of their papers we have counts for
3 papers
Recurrent Value Functions
Pierre Thodoroff, Nishanth Anand, Lucas Caccia +2
Despite recent successes in Reinforcement Learning, value-based methods often suffer from high variance hindering performance. In this paper, we illustrate this in a continuous con…
Temporal Regularization in Markov Decision Process
Pierre Thodoroff, Audrey Durand, Joelle Pineau +1
Several applications of Reinforcement Learning suffer from instability due to high variance. This is especially prevalent in high dimensional domains. Regularization is a commonly…
Adversarial Balancing for Causal Inference
Michal Ozery-Flato, Pierre Thodoroff, Matan Ninio +2
Biases in observational data of treatments pose a major challenge to estimating expected treatment outcomes in different populations. An important technique that accounts for these…