1 paper
Dehong Xu, Liang Qiu, Minseok Kim +2
Pre-trained large-scale language models (LLMs) excel at producing coherent articles, yet their outputs may be untruthful, toxic, or fail to align with user expectations. Current ap…