1 paper
Tianjiao Li, Mengran Yu, Chenyu Shi +6
Large language models (LLMs) possess strong multilingual capabilities, and combining Reinforcement Learning from Human Feedback (RLHF) with translation tasks has shown great potent…