Showing cs.CVShow all
3 papers · 1 filter
cs.CV2024
Schedule Your Edit: A Simple yet Effective Diffusion Noise Schedule for Image Editing
Haonan Lin, Mengmeng Wang, Jiahao Wang +7
Text-guided diffusion models have significantly advanced image editing, enabling high-quality and diverse modifications driven by text prompts. However, effective editing requires…
cs.CV2024
Flipped Classroom: Aligning Teacher Attention with Student in Generalized Category Discovery
Haonan Lin, Wenbin An, Jiahao Wang +6
Recent advancements have shown promise in applying traditional Semi-Supervised Learning strategies to the task of Generalized Category Discovery (GCD). Typically, this involves a t…
cs.CV2024
Knowledge Acquisition Disentanglement for Knowledge-based Visual Question Answering with Large Language Models
Wenbin An, Feng Tian, Jiahao Nie +7
Knowledge-based Visual Question Answering (KVQA) requires both image and world knowledge to answer questions. Current methods first retrieve knowledge from the image and external k…