Showing cs.CLShow all
3 papers · 1 filter
cs.CL2025
Reinforced Lifelong Editing for Language Models
Zherui Li, Houcheng Jiang, Hao Chen +5
Large language models (LLMs) acquire information from pre-training corpora, but their stored knowledge can become inaccurate or outdated over time. Model editing addresses this cha…
cs.CL2025
Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation
Yuanhe Zhang, Fangzhou Xie, Zhenhong Zhou +4
Large Language Diffusion Models (LLDMs) exhibit comparable performance to LLMs while offering distinct advantages in inference speed and mathematical reasoning tasks.The precise an…
cs.CL2024
Speak Out of Turn: Safety Vulnerability of Large Language Models in Multi-turn Dialogue
Zhenhong Zhou, Jiuyang Xiang, Haopeng Chen +3
Large Language Models (LLMs) have been demonstrated to generate illegal or unethical responses, particularly when subjected to "jailbreak." Research on jailbreak has highlighted th…