activity
20232026
most citedConsistent Order Determination of Markov Decision Process

1 citations · 1 across the 2 of their papers we have counts for

collaborators

5 papers

cs.LG2026

Vector-Valued Distributional Reinforcement Learning Policy Evaluation: A Hilbert Space Embedding Approach

Mehrdad Mohammadi, Qi Zheng, Ruoqing Zhu

We propose an (offline) multi-dimensional distributional reinforcement learning framework (KE-DRL) that leverages Hilbert space mappings to estimate the kernel mean embedding of th…

stat.ML2025

Reinforcement Learning with Continuous Actions Under Unmeasured Confounding

Yuhan Li, Eugene Han, Yifan Hu +4

This paper addresses the challenge of offline policy learning in reinforcement learning with continuous action spaces when unmeasured confounders are present. While most existing r…

stat.ME20241 cited

Consistent Order Determination of Markov Decision Process

Chuyun Ye, Lixing Zhu, Ruoqing Zhu

The Markov assumption in Markov Decision Processes (MDPs) is fundamental in reinforcement learning, influencing both theoretical research and practical applications. Existing metho…

stat.ME2023

AI in Pharma for Personalized Sequential Decision-Making: Methods, Applications and Opportunities

Yuhan Li, Hongtao Zhang, Keaven Anderson +2

In the pharmaceutical industry, the use of artificial intelligence (AI) has seen consistent growth over the past decade. This rise is attributed to major advancements in statistica…

stat.ML2023

Stage-Aware Learning for Dynamic Treatments

Hanwen Ye, Wenzhuo Zhou, Ruoqing Zhu +1

Recent advances in dynamic treatment regimes (DTRs) facilitate the search for optimal treatments, which are tailored to individuals' specific needs and able to maximize their expec…