2 papers
cs.CL2025
RIVAL: Reinforcement Learning with Iterative and Adversarial Optimization for Machine Translation
Tianjiao Li, Mengran Yu, Chenyu Shi +6
Large language models (LLMs) possess strong multilingual capabilities, and combining Reinforcement Learning from Human Feedback (RLHF) with translation tasks has shown great potent…
cs.IR2024
An Enhanced Batch Query Architecture in Real-time Recommendation
Qiang Zhang, Zhipeng Teng, Disheng Wu +1
In industrial recommendation systems on websites and apps, it is essential to recall and predict top-n results relevant to user interests from a content pool of billions within mil…