3 citations · 6 across the 4 of their papers we have counts for
4 papers
HelpSteer: Multi-attribute Helpfulness Dataset for SteerLM
Zhilin Wang, Yi Dong, Jiaqi Zeng +8
Existing open-source helpfulness preference datasets do not specify what makes some responses more helpful and others less so. Models trained on these datasets can incidentally lea…
SteerLM: Attribute Conditioned SFT as an (User-Steerable) Alternative to RLHF
Yi Dong, Zhilin Wang, Makesh Narsimhan Sreedhar +2
Model alignment with human preferences is an essential step in making Large Language Models (LLMs) helpful and consistent with human values. It typically consists of supervised fin…
Decentralised and Cooperative Control of Multi-Robot Systems through Distributed Optimisation
Yi Dong, Zhongguo Li, Xingyu Zhao +2
Multi-robot cooperative control has gained extensive research interest due to its wide applications in civil, security, and military domains. This paper proposes a cooperative cont…
Short-term Load Forecasting with Distributed Long Short-Term Memory
Yi Dong, Yang Chen, Xingyu Zhao +1
With the employment of smart meters, massive data on consumer behaviour can be collected by retailers. From the collected data, the retailers may obtain the household profile infor…