1 paper · 1 filter
Zecheng Tang, Keyan Zhou, Juntao Li +5
Text detoxification aims to minimize the risk of language models producing toxic content. Existing detoxification methods of directly constraining the model output or further train…