3 papers
cs.CV2025
BeyondFacial: Identity-Preserving Personalized Generation Beyond Facial Close-ups
Songsong Zhang, Chuanqi Tang, Hongguang Zhang +7
Identity-Preserving Personalized Generation (IPPG) has advanced film production and artistic creation, yet existing approaches overemphasize facial regions, resulting in outputs do…
cs.CV2025
Class-Aware Prototype Learning with Negative Contrast for Test-Time Adaptation of Vision-Language Models
Xiaozhen Qiao, Jingkai Zhao, Yuqiu Jiang +4
Vision-Language Models (VLMs) demonstrate impressive zero-shot generalization through large-scale image-text pretraining, yet their performance can drop once the deployment distrib…
cs.CV2024
MagicNaming: Consistent Identity Generation by Finding a "Name Space" in T2I Diffusion Models
Jing Zhao, Heliang Zheng, Chaoyue Wang +3
Large-scale text-to-image diffusion models, (e.g., DALL-E, SDXL) are capable of generating famous persons by simply referring to their names. Is it possible to make such models gen…