1 paper · 1 filter
Tassilo Klein, Moin Nabi
The generation of toxic content by large language models (LLMs) remains a critical challenge for the safe deployment of language technology. We propose a novel framework for implic…