3 papers
cs.CV2024
SG-Adapter: Enhancing Text-to-Image Generation with Scene Graph Guidance
Guibao Shen, Luozhou Wang, Jiantao Lin +9
Recent advancements in text-to-image generation have been propelled by the development of diffusion models and multi-modality learning. However, since text is typically represented…
cs.CV2023
DVIS++: Improved Decoupled Framework for Universal Video Segmentation
Tao Zhang, Xingye Tian, Yikang Zhou +7
We present the \textbf{D}ecoupled \textbf{VI}deo \textbf{S}egmentation (DVIS) framework, a novel approach for the challenging task of universal video segmentation, including video…
cs.CV2023
Stable Segment Anything Model
Qi Fan, Xin Tao, Lei Ke +6
The Segment Anything Model (SAM) achieves remarkable promptable segmentation given high-quality prompts which, however, often require good skills to specify. To make SAM robust to…