Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
What Makes a Good Reasoning Chain? Uncovering Structural Patterns in Long Chain-of-Thought Reasoning
Gangwei Jiang, Yahui Liu, Zhaoyi Li +5
Recent advances in reasoning with large language models (LLMs) have popularized Long Chain-of-Thought (LCoT), a strategy that encourages deliberate and step-by-step reasoning befor…
cs.AI2024
Refine Large Language Model Fine-tuning via Instruction Vector
Gangwei Jiang, Zhaoyi Li, Defu Lian +1
Fine-tuning large language models (LLMs) can cause them to lose their general capabilities. However, the intrinsic mechanisms behind such forgetting remain unexplored. In this pape…