1 paper
JoonHo Lee, HyeonMin Cho, Jaewoong Yun +3
We present SGuard-v1, a lightweight safety guardrail for Large Language Models (LLMs), which comprises two specialized models to detect harmful content and screen adversarial promp…