1 paper · 1 filter
Mahdi Sabbaghi, Paul Kassianik, George Pappas +3
As large language models (LLMs) are becoming more capable and widespread, the study of their failure cases is becoming increasingly important. Recent advances in standardizing, mea…