2 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.CL2023★ 2 cited
Everyone Deserves A Reward: Learning Customized Human Preferences
Pengyu Cheng, Jiawen Xie, Ke Bai +2
Reward models (RMs) are essential for aligning large language models (LLMs) with human preferences to improve interaction quality. However, the real world is pluralistic, which lea…
cs.CL2023★ 1 cited
Open World Classification with Adaptive Negative Samples
Ke Bai, Guoyin Wang, Jiwei Li +5
Open world classification is a task in natural language processing with key practical relevance and impact. Since the open or {\em unknown} category data only manifests in the infe…