1 paper
Rui Zou, Mengqi Wei, Jintian Feng +3
In recent years, large language models have shown exceptional performance in fulfilling diverse human needs. However, their training data can introduce harmful content, underscorin…