2 papers
cs.CL2026
Denoising Iterative Self-Correction: Structured Verification Loops for Reliable LLM Reasoning
Shen Yin, David Ken, Joel Stremmel
Large language models produce fluent but often incorrect multi-step reasoning, and naive correction methods risk degrading already-correct answers. We introduce Denoising Iterative…
cs.LG2025
Likert or Not: LLM Absolute Relevance Judgments on Fine-Grained Ordinal Scales
Charles Godfrey, Ping Nie, Natalia Ostapuk +3
Large language models (LLMs) obtain state of the art zero shot relevance ranking performance on a variety of information retrieval tasks. The two most common prompts to elicit LLM…