1 paper
Zheyu Lin, Jirui Yang, Yukui Qiu +3
Evaluating the safety robustness of LLMs is critical for their deployment. However, mainstream Red Teaming methods rely on online generation and black-box output analysis. These ap…