5 citations · 5 across the 1 of their papers we have counts for
3 papers · 1 filter
Proportional Response: Contextual Bandits for Simple and Cumulative Regret Minimization
Sanath Kumar Krishnamurthy, Ruohan Zhan, Susan Athey +1
In many applications, e.g. in healthcare and e-commerce, the goal of a contextual bandit may be to learn an optimal treatment assignment policy at the end of the experiment. That i…
Constrained Reinforcement Learning for Short Video Recommendation
Qingpeng Cai, Ruohan Zhan, Chi Zhang +5
The wide popularity of short videos on social media poses new opportunities and challenges to optimize recommender systems on the video-sharing platforms. Users provide complex and…
Towards Content Provider Aware Recommender Systems: A Simulation Study on the Interplay between User and Provider Utilities
Ruohan Zhan, Konstantina Christakopoulou, Ya Le +6
Most existing recommender systems focus primarily on matching users to content which maximizes user satisfaction on the platform. It is increasingly obvious, however, that content…