1 paper · 1 filter
Ming Li, Yanhong Li, Ziyue Li +1
As the post-training of large language models (LLMs) advances from instruction-following to complex reasoning tasks, understanding how different data affect finetuning dynamics rem…