Showing stat.MLShow all
2 papers · 1 filter
stat.ML2026
Reliable Chain-of-Thought via Prefix Consistency
Naoto Iwase, Yuki Ichihara, Mohammad Atif Quamar +1
Large Language Models often improve accuracy on reasoning tasks by sampling multiple Chain-of-Thought (CoT) traces and aggregating them with majority voting (MV), a test-time techn…
stat.ML2026
CITE: Anytime-Valid Statistical Inference in LLM Self-Consistency
Hirofumi Ota, Naoto Iwase, Yuki Ichihara +2
Large language models often improve reasoning by sampling multiple outputs and aggregating their final answers, but precise and efficient control of error levels remains a challeng…