1 paper · 1 filter
Xuan Luo, Yubin Chen, Zhiyu Hou +4
As large language models (LLMs) are increasingly embedded in everyday decision-making, their safety responsibilities extend beyond reacting to explicit harmful intent toward antici…