1 paper
Fenghua Weng, Jian Lou, Jun Feng +2
Safety alignment is critical in pre-training large language models (LLMs) to generate responses aligned with human values and refuse harmful queries. Unlike LLM, the current safety…