1 paper · 1 filter
Geng Liu, Li Feng, Carlo Alberto Bono +3
Recent research has highlighted that assigning specific personas to large language models (LLMs) can significantly increase harmful content generation. However, limited attention h…