2 papers
cs.CR2026
Hidden in Thought: Transferable Chain-of-Thought Artifacts Induce Harmful Behavior
Ali khalil, Aly M. Kassem, Mohamed Abdelrazek +3
We investigate whether harmful chain-of-thought (CoT) traces from compromised language models can transfer unsafe behaviour and be distilled into reusable jailbreak attacks. Using…
cs.SE2025
Ensuring Robustness in ML-enabled Software Systems: A User Survey
Hala Abdelkader, Mohamed Abdelrazek, Priya Rani +2
Ensuring robustness in ML-enabled software systems requires addressing critical challenges, such as silent failures, out-of-distribution (OOD) data, and adversarial attacks. Tradit…