3 papers
cs.CL2026
Relax: An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale
Liujie Zhang, Benzhe Ning, Rui Yang +8
Reinforcement learning (RL) post-training has proven effective at unlocking reasoning, self-reflection, and tool-use capabilities in large language models. As models extend to omni…
cs.IR2025
Towards Large-scale Generative Ranking
Yanhua Huang, Yuqi Chen, Xiong Cao +17
Generative recommendation has recently emerged as a promising paradigm in information retrieval. However, generative ranking systems are still understudied, particularly with respe…
cs.LG2024
Optimizing Personalized Federated Learning through Adaptive Layer-Wise Learning
Weihang Chen, Cheng Yang, Jie Ren +2
Real-life deployment of federated Learning (FL) often faces non-IID data, which leads to poor accuracy and slow convergence. Personalized FL (pFL) tackles these issues by tailoring…