Showing cs.CVShow all
2 papers · 1 filter
cs.CV2024
Multi-Class Textual-Inversion Secretly Yields a Semantic-Agnostic Classifier
Kai Wang, Fei Yang, Bogdan Raducanu +1
With the advent of large pre-trained vision-language models such as CLIP, prompt learning methods aim to enhance the transferability of the CLIP model. They learn the prompt given…
cs.CV2024
LocInv: Localization-aware Inversion for Text-Guided Image Editing
Chuanming Tang, Kai Wang, Fei Yang +1
Large-scale Text-to-Image (T2I) diffusion models demonstrate significant generation capabilities based on textual prompts. Based on the T2I diffusion models, text-guided image edit…