1 paper · 1 filter
Rongwu Xu, Yishuo Cai, Zhenhong Zhou +6
The risk of harmful content generated by large language models (LLMs) becomes a critical concern. This paper presents a systematic study on assessing and improving LLMs' capability…