2 citations · 2 across the 1 of their papers we have counts for
1 paper · 1 filter
Lize Alberts, Benjamin Ellis, Andrei Lupu +1
We introduce a multi-turn benchmark for evaluating personalised alignment in LLM-based AI assistants, focusing on their ability to handle user-provided safety-critical contexts. Ou…