2 papers
cs.CL2025
In-context Demonstration Matters: On Prompt Optimization for Pseudo-Supervision Refinement
Zhen-Yu Zhang, Jiandong Zhang, Huaxiu Yao +2
Large language models (LLMs) have achieved great success across diverse tasks, and fine-tuning is sometimes needed to further enhance generation quality. Most existing methods rely…
cs.LG2024
Generating Chain-of-Thoughts with a Pairwise-Comparison Approach to Searching for the Most Promising Intermediate Thought
Zhen-Yu Zhang, Siwei Han, Huaxiu Yao +2
To improve the ability of the large language model (LLMs) to tackle complex reasoning problems, chain-of-thoughts (CoT) methods were proposed to guide LLMs to reason step-by-step,…