1 paper · 1 filter
Lingkai Kong, Haorui Wang, Wenhao Mu +7
Aligning large language models (LLMs) with human objectives is crucial for real-world applications. However, fine-tuning LLMs for alignment often suffers from unstable training and…