1 citations · 1 across the 1 of their papers we have counts for
1 paper
Maximilian Mozes, Jessica Hoffmann, Katrin Tomanek +5
Text-based safety classifiers are widely used for content moderation and increasingly to tune generative language model behavior - a topic of growing concern for the safety of digi…