3 papers
cs.CY2025
Reproducible workflow for online AI in digital health
Susobhan Ghosh, Bhanu T. Gullapalli, Daiqi Gao +5
Online artificial intelligence (AI) algorithms are an important component of digital health interventions. These online algorithms are designed to continually learn and improve the…
cs.LG2025
Active Measuring in Reinforcement Learning With Delayed Negative Effects
Daiqi Gao, Ziping Xu, Aseel Rawashdeh +2
Measuring states in reinforcement learning (RL) can be costly in real-world settings and may negatively influence future outcomes. We introduce the Actively Observable Markov Decis…
math.ST2025
Statistical Inference for Misspecified Contextual Bandits
Yongyi Guo, Ziping Xu
Contextual bandit algorithms have transformed modern experimentation by enabling real-time adaptation for personalized treatment and efficient use of data. Yet these advantages cre…