2 papers
cs.CL2024
REInstruct: Building Instruction Data from Unlabeled Corpus
Shu Chen, Xinyan Guan, Yaojie Lu +3
Manually annotating instruction data for large language models is difficult, costly, and hard to scale. Meanwhile, current automatic annotation methods typically rely on distilling…
cs.CL2024
SoFA: Shielded On-the-fly Alignment via Priority Rule Following
Xinyu Lu, Bowen Yu, Yaojie Lu +5
The alignment problem in Large Language Models (LLMs) involves adapting them to the broad spectrum of human values. This requirement challenges existing alignment methods due to di…