1 paper
Xia Hou, Qifeng Li, Jian Yang +8
Instruction tuning as an effective technique aligns the outputs of large language models (LLMs) with human preference. But how to generate the seasonal multi-turn dialogues from ra…