Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
StarFT: Robust Fine-tuning of Zero-shot Models via Spuriosity Alignment
Younghyun Kim, Jongheon Jeong, Sangkyung Kwak +3
Learning robust representations from data often requires scale, which has led to the success of recent zero-shot models such as CLIP. However, the obtained robustness can easily be…
cs.AI2024
Safety Alignment Backfires: Preventing the Re-emergence of Suppressed Concepts in Fine-tuned Text-to-Image Diffusion Models
Sanghyun Kim, Moonseok Choi, Jinwoo Shin +1
Fine-tuning text-to-image diffusion models is widely used for personalization and adaptation for new domains. In this paper, we identify a critical vulnerability of fine-tuning: sa…