3 papers
cs.LG2025
Group-Sensitive Offline Contextual Bandits
Yihong Guo, Junjie Luo, Guodong Gao +2
Offline contextual bandits allow one to learn policies from historical/offline data without requiring online interaction. However, offline policy optimization that maximizes overal…
cs.AI2025
PAME-AI: Patient Messaging Creation and Optimization using Agentic AI
Junjie Luo, Yihong Guo, Anqi Liu +2
Messaging patients is a critical part of healthcare communication, helping to improve things like medication adherence and healthy behaviors. However, traditional mobile message de…
cs.LG2024
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation
Yihong Guo, Yixuan Wang, Yuanyuan Shi +2
Training a policy in a source domain for deployment in the target domain under a dynamics shift can be challenging, often resulting in performance degradation. Previous work tackle…