1 paper
Aleksei Ilin, Gor Matevosyan, Xueying Ma +5
We introduce a lightweight yet highly effective safety guardrail framework for language models, demonstrating that small-scale language models can achieve, and even surpass, the pe…