2 papers
cs.IR2025
Value Function Decomposition in Markov Recommendation Process
Xiaobei Wang, Shuchang Liu, Qingpeng Cai +4
Recent advances in recommender systems have shown that user-system interaction essentially formulates long-term optimization problems, and online reinforcement learning can be adop…
cs.IR2024
Future Impact Decomposition in Request-level Recommendations
Xiaobei Wang, Shuchang Liu, Xueliang Wang +6
In recommender systems, reinforcement learning solutions have shown promising results in optimizing the interaction sequence between users and the system over the long-term perform…