159 citations · 628 across the 32 of their papers we have counts for
4 papers · 1 filter
Factual and Personalized Recommendations using Language Models and Reinforcement Learning
Jihwan Jeong, Yinlam Chow, Guy Tennenholtz +4
Recommender systems (RSs) play a central role in connecting users to content, products, and services, matching candidate items to users based on their preferences. While traditiona…
Policy-Aware Model Learning for Policy Gradient Methods
Romina Abachi, Mohammad Ghavamzadeh, Amir-massoud Farahmand
This paper considers the problem of learning a model in model-based reinforcement learning (MBRL). We examine how the planning module of an MBRL algorithm uses the model, and propo…
Path Consistency Learning in Tsallis Entropy Regularized MDPs
Ofir Nachum, Yinlam Chow, Mohammad Ghavamzadeh
We study the sparse entropy-regularized reinforcement learning (ERL) problem in which the entropy term is a special form of the Tsallis entropy. The optimal policy of this formulat…
Policy Gradient for Coherent Risk Measures
Aviv Tamar, Yinlam Chow, Mohammad Ghavamzadeh +1
Several authors have recently developed risk-sensitive policy gradient methods that augment the standard expected cost minimization problem with a measure of variability in cost. T…