1 paper
Seorin Kim, Dongyoung Lee, Jaejin Lee
Large language models (LLMs) often exhibit societal biases in their outputs, prompting ethical concerns regarding fairness and harm. In this work, we propose KLAAD (KL-Attention Al…