Showing stat.MLShow all
2 papers · 1 filter
stat.ML2026
Improving Backward Conformal Prediction via Non-Conformity Score Transformation
Junxian Liu, Hao Zeng, Hongxin Wei
Conformal Prediction (CP) provides a statistical framework for uncertainty quantification that constructs prediction sets with coverage guarantees. While CP yields uncontrolled pre…
stat.ML2026
An Interpretable and Scalable Framework for Evaluating Large Language Models
Xinhao Qu, Qiang Heng, Hao Zeng +1
Evaluation of large language models (LLMs) is increasingly critical, yet standard benchmarking methods rely on average accuracy, overlooking both the inherent stochasticity of LLM…