3 papers
cs.LG2025
Active Measuring in Reinforcement Learning With Delayed Negative Effects
Daiqi Gao, Ziping Xu, Aseel Rawashdeh +2
Measuring states in reinforcement learning (RL) can be costly in real-world settings and may negatively influence future outcomes. We introduce the Actively Observable Markov Decis…
stat.ML2025
Counterfactual inference in sequential experiments
Raaz Dwivedi, Katherine Tian, Sabina Tomkins +3
We consider after-study statistical inference for sequentially designed experiments wherein multiple units are assigned treatments for multiple time points using treatment policies…
cs.LG2025
Harnessing Causality in Reinforcement Learning With Bagged Decision Times
Daiqi Gao, Hsin-Yu Lai, Predrag Klasnja +1
We consider reinforcement learning (RL) for a class of problems with bagged decision times. A bag contains a finite sequence of consecutive decision times. The transition dynamics…