3 papers
cs.LG2026
SAGE: Sequence-level Adaptive Gradient Evolution for Generative Recommendation
Yu Xie, Xing Kai Ren, Ying Qi +1
Reinforcement learning-based preference optimization is increasingly used to align list-wise generative recommenders with complex, multi-objective user feedback, yet existing optim…
cs.AI2025
RecLLM-R1: A Two-Stage Training Paradigm with Reinforcement Learning and Chain-of-Thought v1
Yu Xie, Xingkai Ren, Ying Qi +2
Traditional recommendation systems often grapple with "filter bubbles", underutilization of external knowledge, and a disconnect between model optimization and business policy iter…
cs.IR2024
LLM4PR: Improving Post-Ranking in Search Engine with Large Language Models
Yang Yan, Yihao Wang, Chi Zhang +8
Alongside the rapid development of Large Language Models (LLMs), there has been a notable increase in efforts to integrate LLM techniques in information retrieval (IR) and search e…