3 papers
cs.IR2025
R1-Ranker: Teaching LLM Rankers to Reason
Tao Feng, Zhigang Hua, Zijie Lei +4
Large language models (LLMs) have recently shown strong reasoning abilities in domains like mathematics, coding, and scientific problem-solving, yet their potential for ranking tas…
cs.IR2025
Unified Semantic and ID Representation Learning for Deep Recommenders
Guanyu Lin, Zhigang Hua, Tao Feng +3
Effective recommendation is crucial for large-scale online platforms. Traditional recommendation systems primarily rely on ID tokens to uniquely identify items, which can effective…
cs.LG2025
Session-Level Dynamic Ad Load Optimization using Offline Robust Reinforcement Learning
Tao Liu, Qi Xu, Wei Shi +2
Session-level dynamic ad load optimization aims to personalize the density and types of delivered advertisements in real time during a user's online session by dynamically balancin…