Showing cs.CRShow all
3 papers · 1 filter
cs.CR2026
GUARD-SLM: Token Activation-Based Defense Against Jailbreak Attacks for Small Language Models
Md Jueal Mia, Joaquin Molto, Yanzhao Wu +1
Small Language Models (SLMs) are emerging as efficient and economically viable alternatives to Large Language Models (LLMs), offering competitive performance with significantly low…
cs.CR2026
Multi-turn Jailbreaking Attack in Multi-Modal Large Language Models
Badhan Chandra Das, Md Tasnim Jawad, Joaquin Molto +2
In recent years, the security vulnerabilities of Multi-modal Large Language Models (MLLMs) have become a serious concern in the Generative Artificial Intelligence (GenAI) research.…
cs.CR2025
System Prompt Extraction Attacks and Defenses in Large Language Models
Badhan Chandra Das, M. Hadi Amini, Yanzhao Wu
The system prompt in Large Language Models (LLMs) plays a pivotal role in guiding model behavior and response generation. Often containing private configuration details, user roles…