23 citations · 23 across the 2 of their papers we have counts for
2 papers
cs.CL2024
Regularizing Hidden States Enables Learning Generalizable Reward Model for LLMs
Rui Yang, Ruomeng Ding, Yong Lin +2
Reward models trained on human preference data have been proven to effectively align Large Language Models (LLMs) with human intent within the framework of reinforcement learning f…
cond-mat.stat-mech2016★ 23 cited
Multifractality and Laplace spectrum of horizontal visibility graphs constructed from fractional Brownian motions
Zu-Guo Yu, Huan Zhang, Da-Wen Huang +2
Many studies have shown that additional information can be gained on time series by investigating their associated complex networks. In this work, we investigate the multifractal p…