16 citations · 26 across the 3 of their papers we have counts for
3 papers
cs.CL2024★ 2 cited
Advancing Translation Preference Modeling with RLHF: A Step Towards Cost-Effective Solution
Nuo Xu, Jun Zhao, Can Zu +9
Faithfulness, expressiveness, and elegance is the constant pursuit in machine translation. However, traditional metrics like \textit{BLEU} do not strictly align with human preferen…
cs.AI2024★ 8 cited
Secrets of RLHF in Large Language Models Part II: Reward Modeling
Binghai Wang, Rui Zheng, Lu Chen +24
Reinforcement Learning from Human Feedback (RLHF) has become a crucial technology for aligning language models with human values and intentions, enabling models to produce more hel…
cs.RO2014★ 16 cited
GP-Localize: Persistent Mobile Robot Localization using Online Sparse Gaussian Process Observation Model
Nuo Xu, Kian Hsiang Low, Jie Chen +2
Central to robot exploration and mapping is the task of persistent localization in environmental fields characterized by spatially correlated measurements. This paper presents a Ga…