1 paper · 1 filter
Lisong Sun, Li Wang, Chen Zhang +4
Large language models (LLMs) have achieved remarkable progress, with post-training playing a crucial role in enhancing their reasoning capabilities. Among post-training paradigms,…