From the 1 of 2 linked papers with an AI index.
2 papers
cs.LG2026
Random Logit Scaling: Defending Deep Neural Networks Against Black-Box Score-Based Adversarial Example Attacks
Hamid Dashtbani, Mehdi Dousti Gandomani, AmirMahdi Sadeghzadeh
The paper introduces Random Logit Scaling, a plug‑and‑play post‑processing defense that randomly rescales model logits to thwart black‑box score‑based adversarial attacks while kee…
cs.LG2025
TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts
Torsten KrauÃ, Hamid Dashtbani, Alexandra Dmitrienko
Machine learning is advancing rapidly, with applications bringing notable benefits, such as improvements in translation and code generation. Models like ChatGPT, powered by Large L…