1 paper
Baolong Bi, Shaohan Huang, Yiwei Wang +11
Reliable responses from large language models (LLMs) require adherence to user instructions and retrieved information. While alignment techniques help LLMs align with human intenti…