3 papers
cs.GT2026
Simple and Robust Quality Disclosure: The Power of Quantile Partition
Shipra Agrawal, Yiding Feng, Wei Tang
Quality information on online platforms is often conveyed through simple, percentile-based badges and tiers that remain stable across different market environments. Motivated by th…
stat.ML2025
Reinforcement Learning in MDPs with Information-Ordered Policies
Zhongjun Zhang, Shipra Agrawal, Ilan Lobel +2
We propose an epoch-based reinforcement learning algorithm for infinite-horizon average-cost Markov decision processes (MDPs) that leverages a partial order over a policy class. In…
cs.LG2025
Q-learning with Posterior Sampling
Priyank Agrawal, Shipra Agrawal, Azmat Azati
Bayesian posterior sampling techniques have demonstrated superior empirical performance in many exploration-exploitation settings. However, their theoretical analysis remains a cha…