Showing cs.LGShow all
3 papers · 1 filter
cs.LG2024
A Counterfactual Analysis of the Dishonest Casino
Martin Haugh, Raghav Singal
The dishonest casino is a well-known hidden Markov model (HMM) often used in education to introduce HMMs and graphical models. A sequence of die rolls is observed with the casino s…
cs.LG2024
Model-Free Approximate Bayesian Learning for Large-Scale Conversion Funnel Optimization
Garud Iyengar, Raghav Singal
The flexibility of choosing the ad action as a function of the consumer state is critical for modern-day marketing campaigns. We study the problem of identifying the optimal sequen…
cs.LG2018
A Finite Time Analysis of Temporal Difference Learning With Linear Function Approximation
Jalaj Bhandari, Daniel Russo, Raghav Singal
Temporal difference learning (TD) is a simple iterative algorithm used to estimate the value function corresponding to a given policy in a Markov decision process. Although TD is o…