2 citations · 2 across the 3 of their papers we have counts for
3 papers
cs.CL2024
CSS: Contrastive Semantic Similarity for Uncertainty Quantification of LLMs
Shuang Ao, Stefan Rueger, Advaith Siddharthan
Despite the impressive capability of large language models (LLMs), knowing when to trust their generations remains an open challenge. The recent literature on uncertainty quantific…
cs.AI2023
Empirical Optimal Risk to Quantify Model Trustworthiness for Failure Detection
Shuang Ao, Stefan Rueger, Advaith Siddharthan
Failure detection (FD) in AI systems is a crucial safeguard for the deployment for safety-critical tasks. The common evaluation method of FD performance is the Risk-coverage (RC) c…
cs.LG2023★ 2 cited
Two Sides of Miscalibration: Identifying Over and Under-Confidence Prediction for Network Calibration
Shuang Ao, Stefan Rueger, Advaith Siddharthan
Proper confidence calibration of deep neural networks is essential for reliable predictions in safety-critical tasks. Miscalibration can lead to model over-confidence and/or under-…