2 papers
cs.CL2025
Uncertainty Quantification and Confidence Calibration in Large Language Models: A Survey
Xiaoou Liu, Tiejin Chen, Longchao Da +3
Large Language Models (LLMs) excel in text generation, reasoning, and decision-making, enabling their adoption in high-stakes domains such as healthcare, law, and transportation. H…
cs.CL2025
MCQA-Eval: Efficient Confidence Evaluation in NLG with Gold-Standard Correctness Labels
Xiaoou Liu, Zhen Lin, Longchao Da +3
Large Language Models (LLMs) require robust confidence estimation, particularly in critical domains like healthcare and law where unreliable outputs can lead to significant consequ…