Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
Do Methods to Jailbreak and Defend LLMs Generalize Across Languages?
Berk Atil, Rebecca J. Passonneau, Fred Morstatter
Large language models (LLMs) undergo safety alignment after training and tuning, yet recent work shows that safety can be bypassed through jailbreak attacks. While many jailbreaks…
cs.CL2025
Knowledge Graph Analysis of Legal Understanding and Violations in LLMs
Abha Jha, Abel Salinas, Fred Morstatter
The rise of Large Language Models (LLMs) offers transformative potential for interpreting complex legal frameworks, such as Title 18 Section 175 of the US Code, which governs biolo…