1 paper · 1 filter
Ilias Chalkidis, Anders Søgaard
The scope of AI safety and alignment work in generative artificial intelligence (GenAI) has so far mostly been limited to harms related to: (a) discrimination and hate speech, (b)…