1 paper
Zecheng Tang, Keyan Zhou, Juntao Li +5
Text detoxification aims to minimize the risk of language models producing toxic content. Existing detoxification methods of directly constraining the model output or further train…