4 papers
Re-evaluating Confidence Remasking in Masked Diffusion Language Models
Stipe Frkovic, Metod Jazbec, Dan Zhang +3
Masked diffusion language models (dLLMs) have recently emerged as a competitive alternative to autoregressive language models, with the promise of faster inference via parallel tok…
Human-AI Teaming Through the Lens of Calibration
Eric Nalisnick, Chi Zhang, Sophia Qian +1
We study models for human-AI teaming through the lens of statistical calibration. We assume the team consists of an AI model and human -- both of which are calibrated with respect…
On Continuous Monitoring of Risk Violations under Unknown Shift
Alexander Timans, Rajeev Verma, Eric Nalisnick +1
Machine learning systems deployed in the real world must operate under dynamic and often unpredictable distribution shifts. This challenges the validity of statistical safety assur…
On Calibration in Multi-Distribution Learning
Rajeev Verma, Volker Fischer, Eric Nalisnick
Modern challenges of robustness, fairness, and decision-making in machine learning have led to the formulation of multi-distribution learning (MDL) frameworks in which a predictor…