2 papers
cs.LG2025
Reinforcement Learning from Human Feedback with High-Confidence Safety Constraints
Yaswanth Chittepu, Blossom Metevier, Will Schwarzer +3
Existing approaches to language model alignment often treat safety as a tradeoff against helpfulness, which can lead to unacceptable responses in sensitive domains. To ensure relia…
cs.SD2025
Are Deep Speech Denoising Models Robust to Adversarial Noise?
Will Schwarzer, Neel Chaudhari, Philip S. Thomas +2
Deep noise suppression (DNS) models enjoy widespread use throughout a variety of high-stakes speech applications. However, we show that four recent DNS models can each be reduced t…