1 paper
Bowen Zhang, Meiyi Wang, Harold Soh
Post-training improves instruction-following and helpfulness of large language models (LLMs) but often reduces generation diversity, which leads to repetitive outputs in open-ended…