1 paper · 1 filter
Bartosz Dziuba, Kacper Kuchta, PaweÅ Batorski +2
Large Language Models (LLMs) have improved substantially alignment, yet their behavior remains highly sensitive to prompt phrasing. This brittleness has motivated automated prompt…