2 papers
cs.CL2025
QA-Calibration of Language Model Confidence Scores
Putra Manggala, Atalanti Mastakouri, Elke Kirschbaum +2
To use generative question-and-answering (QA) systems for decision-making and in any critical application, these systems need to provide well-calibrated confidence scores that refl…
cs.LG2024
Learning to Defer to a Population: A Meta-Learning Approach
Dharmesh Tailor, Aditya Patra, Rajeev Verma +2
The learning to defer (L2D) framework allows autonomous systems to be safe and robust by allocating difficult decisions to a human expert. All existing work on L2D assumes that eac…