Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
Unforgotten Safety: Preserving Safety Alignment of Large Language Models with Continual Learning
Lama Alssum, Hani Itani, Hasan Abed Al Kader Hammoud +3
The safety alignment of large language models (LLMs) is becoming increasingly important with their democratization. In this paper, we study the safety degradation that comes with a…
cs.CL2025
Beyond the Last Answer: Your Reasoning Trace Uncovers More than You Think
Hasan Abed Al Kader Hammoud, Hani Itani, Bernard Ghanem
Large Language Models (LLMs) leverage step-by-step reasoning to solve complex problems. Standard evaluation practice involves generating a complete reasoning trace and assessing th…