2 papers
cs.CL2025
Confidence-Weighted Token Set Cover for Early Hypothesis Pruning in Self-Consistency
Md Arafat Sultan, Ramón Fernandez Astudillo
Despite its simplicity and efficacy, the high token expenditure of self-consistency can limit its practical utility. Here we investigate if self-consistency can be made more token-…
cs.LG2025
Optimal Policy Minimum Bayesian Risk
Ramón Fernandez Astudillo, Md Arafat Sultan, Aashka Trivedi +4
Inference scaling helps LLMs solve complex reasoning problems through extended runtime computation. On top of long chain-of-thought (long-CoT) models, purely inference-time techniq…