6 citations · 6 across the 1 of their papers we have counts for
1 paper · 1 filter
Yi Zeng, Hongpeng Lin, Jingwen Zhang +3
Most traditional AI safety research has approached AI models as machines and centered on algorithm-focused attacks developed by security experts. As large language models (LLMs) be…