2 papers
cs.LG2026
Jailbreak Scaling Laws for Large Language Models: Polynomial-Exponential Crossover
Indranil Halder, Annesya Banerjee, Cengiz Pehlevan
Adversarial attacks can reliably steer safety-aligned large language models toward unsafe behavior. Empirically, we find that adversarial prompt-injection attacks can amplify attac…
cs.SD2024
Incorporating Talker Identity Aids With Improving Speech Recognition in Adversarial Environments
Sagarika Alavilli, Annesya Banerjee, Gasser Elbanna +1
Current state-of-the-art speech recognition models are trained to map acoustic signals into sub-lexical units. While these models demonstrate superior performance, they remain vuln…