2 papers
cs.IR2026
A Self-Triggered Agentic Push Recommendation System
Zhao-Yu Zhang, Qingying Chen, Chunyuan Zheng +9
Push notification is a critical recommendation scenario on large-scale platforms, allowing the system to proactively reach users outside the application to improve long-term re-eng…
cs.AI2026
QLPO: Quadrant-weighted Sampling for Length-aware Policy Optimization
Siwei Chen, Siqi Chen, Xupeng Miao +1
Recent large reasoning models often develop long chain-of-thought responses during reinforcement learning (RL), resulting in high inference latency and deployment cost. Existing me…