1 paper · 1 filter
Minchan Kwon, Sunghyun Baek, Minseo Kim +3
Large Language Model (LLM) Red-Teaming, which proactively identifies vulnerabilities of LLMs, is an essential process for ensuring safety. Finding effective and diverse attacks in…