4 citations · 5 across the 41 of their papers we have counts for
1 paper · 1 filter
Amr Hegazy, Mostafa Elhoushi, Amr Alanwar
Controlling undesirable Large Language Model (LLM) behaviors, such as the generation of unsafe content or failing to adhere to safety guidelines, often relies on costly fine-tuning…