From the 1 of 3 linked papers with an AI index.
1 paper · 1 filter
Paul Darm, Annalisa Riccardi
Robust alignment guardrails for large language models (LLMs) are becoming increasingly important with their widespread application. In contrast to previous studies, we demonstrate…