1 paper · 1 filter
Jin Zhang, Flood Sung, Zhilin Yang +2
In the field of large language model (LLM) post-training, the effectiveness of utilizing synthetic data generated by the LLM itself has been well-presented. However, a key question…