2 citations · 2 across the 2 of their papers we have counts for
2 papers
stat.ML2024
Causal Deepsets for Off-policy Evaluation under Spatial or Spatio-temporal Interferences
Runpeng Dai, Jianing Wang, Fan Zhou +4
Off-policy evaluation (OPE) is widely applied in sectors such as pharmaceuticals and e-commerce to evaluate the efficacy of novel products or policies from offline datasets. This p…
cs.LG2023★ 2 cited
A Better Match for Drivers and Riders: Reinforcement Learning at Lyft
Xabi Azagirre, Akshay Balwally, Guillaume Candeli +16
To better match drivers to riders in our ridesharing application, we revised Lyft's core matching algorithm. We use a novel online reinforcement learning approach that estimates th…