4 citations · 4 across the 2 of their papers we have counts for
3 papers
cs.CY2026
LSSF: Safety Alignment for Large Language Models through Low-Rank Safety Subspace Fusion
Guanghao Zhou, Panjia Qiu, Cen Chen +4
The safety mechanisms of large language models (LLMs) exhibit notable fragility, as even fine-tuning on datasets without harmful content may still undermine their safety capabiliti…
cs.CR2025
CCJA: Context-Coherent Jailbreak Attack for Aligned Large Language Models
Guanghao Zhou, Panjia Qiu, Mingyuan Fan +4
Despite explicit alignment efforts for large language models (LLMs), they can still be exploited to trigger unintended behaviors, a phenomenon known as "jailbreaking." Current jail…
cs.CL2024★ 4 cited
From Beginner to Expert: Modeling Medical Knowledge into General LLMs
Qiang Li, Xiaoyan Yang, Haowen Wang +14
Recently, large language model (LLM) based artificial intelligence (AI) systems have demonstrated remarkable capabilities in natural language understanding and generation. However,…