1 paper
Fan Liu, Yue Feng, Zhao Xu +4
Despite advancements in enhancing LLM safety against jailbreak attacks, evaluating LLM defenses remains a challenge, with current methods often lacking explainability and generaliz…