Showing 2025Show all
2 papers · 1 filter
cs.CV2025
SELFI: Selective Fusion of Identity for Generalizable Deepfake Detection
Younghun Kim, Minsuk Jang, Myung-Joon Kwon +2
Face identity provides a powerful signal for deepfake detection. Prior studies show that even when not explicitly modeled, classifiers often learn identity features implicitly. Thi…
cs.CV2025
Benign-to-Toxic Jailbreaking: Inducing Harmful Responses from Harmless Prompts
Hee-Seon Kim, Minbeom Kim, Wonjun Lee +2
Optimization-based jailbreaks typically adopt the Toxic-Continuation setting in large vision-language models (LVLMs), following the standard next-token prediction objective. In thi…