17 citations · 19 across the 4 of their papers we have counts for
4 papers
Selective Safety Steering via Value-Filtered Decoding
Bat-Sheva Einbinder, Hen Davidov, Yee Whye Teh +2
While large language models (LLMs) are trained to align with human values, their generations may still violate safety constraints. A growing line of work addresses this problem by…
Semi-Supervised Risk Control via Prediction-Powered Inference
Bat-Sheva Einbinder, Liran Ringel, Yaniv Romano
The risk-controlling prediction sets (RCPS) framework is a general tool for transforming the output of any machine learning model to design a predictive rule with rigorous error ra…
Label Noise Robustness of Conformal Prediction
Bat-Sheva Einbinder, Shai Feldman, Stephen Bates +3
We study the robustness of conformal prediction, a powerful tool for uncertainty quantification, to label noise. Our analysis tackles both regression and classification problems, c…
Training Uncertainty-Aware Classifiers with Conformalized Deep Learning
Bat-Sheva Einbinder, Yaniv Romano, Matteo Sesia +1
Deep neural networks are powerful tools to detect hidden patterns in data and leverage them to make predictions, but they are not designed to understand uncertainty and estimate re…