1 paper
Zhexin Zhang, Jiale Cheng, Hao Sun +5
Large pretrained language models can easily produce toxic or biased content, which is prohibitive for practical use. In order to detect such toxic generations, existing methods rel…