1 paper
Yi Dong, Ronghui Mu, Yanghao Zhang +9
In the burgeoning field of Large Language Models (LLMs), developing a robust safety mechanism, colloquially known as "safeguards" or "guardrails", has become imperative to ensure t…