1 paper
Trung Quoc Luong, Xinbo Zhang, Zhanming Jie +3
One way to enhance the reasoning capability of Large Language Models (LLMs) is to conduct Supervised Fine-Tuning (SFT) using Chain-of-Thought (CoT) annotations. This approach does…