2 papers
cs.AI2026
Decoding ML Decision: An Agentic Reasoning Framework for Large-Scale Ranking System
Longfei Yun, Yihan Wu, Haoran Liu +11
Modern large-scale ranking systems operate within a sophisticated landscape of competing objectives, operational constraints, and evolving product requirements. Progress in this do…
cs.IR2026
Objective Shaping with Hard Negatives: Windowed Partial AUC Optimization for RL-based LLM Recommenders
Wentao Shi, Qifan Wang, Chen Chen +7
Reinforcement learning (RL) effectively optimizes Large Language Model (LLM)-based recommenders by contrasting positive and negative items. Empirically, training with beam-search n…