1 paper
Lisong Sun, Li Wang, Chen Zhang +4
Large language models (LLMs) have achieved remarkable progress, with post-training playing a crucial role in enhancing their reasoning capabilities. Among post-training paradigms,…