3 papers
cs.CL2026
What Intermediate Layers Know: Detecting Jailbreaks from Entropy Dynamics
Sofiia Nikolenko, Michele Papucci, Mina Rezaei +1
Jailbreak attacks reveal a persistent weakness in aligned Large Language Models: carefully crafted prompts can elicit policy-violating responses despite safety training. While most…
stat.AP2026
Predicting the 2026 FIFA World Cup with Sufficient Dimension Reduction of Elo Rating Histories
Mina Rezaei, S. Yaser Samadi
We study probabilistic forecasting of the 2026 FIFA World Cup, the first edition with 48 teams and an added Round of 32. The main idea is to describe team strength not only by the…
cs.CL2025
Mitigating Catastrophic Forgetting in Continual Learning through Model Growth
Ege Süalp, Mina Rezaei
Catastrophic forgetting is a significant challenge in continual learning, in which a model loses prior knowledge when it is fine-tuned on new tasks. This problem is particularly cr…