1 paper
McNair Shah, Saleena Angeline, Adhitya Rajendra Kumar +5
Recent advances in large language models (LLMs) have intensified the need to understand and reliably curb their harmful behaviours. We introduce a multidimensional framework for pr…