5 citations · 6 across the 3 of their papers we have counts for
4 papers
Benign-to-Toxic Jailbreaking: Inducing Harmful Responses from Harmless Prompts
Hee-Seon Kim, Minbeom Kim, Wonjun Lee +2
Optimization-based jailbreaks typically adopt the Toxic-Continuation setting in large vision-language models (LVLMs), following the standard next-token prediction objective. In thi…
Parameter Efficient Mamba Tuning via Projector-targeted Diagonal-centric Linear Transformation
Seokil Ham, Hee-Seon Kim, Sangmin Woo +1
Despite the growing interest in Mamba architecture as a potential replacement for Transformer architecture, parameter-efficient fine-tuning (PEFT) approaches for Mamba remain large…
Improving the Transferability of Targeted Adversarial Examples through Object-Based Diverse Input
Junyoung Byun, Seungju Cho, Myung-Joon Kwon +2
The transferability of adversarial examples allows the deception on black-box models, and transfer-based targeted attacks have attracted a lot of interest due to their practical ap…
Balancing Domain Experts for Long-Tailed Camera-Trap Recognition
Byeongjun Park, Jeongsoo Kim, Seungju Cho +2
Label distributions in camera-trap images are highly imbalanced and long-tailed, resulting in neural networks tending to be biased towards head-classes that appear frequently. Alth…