94 citations · 95 across the 2 of their papers we have counts for
7 papers · 1 filter
DiffI2I: Efficient Diffusion Model for Image-to-Image Translation
Bin Xia, Yulun Zhang, Shiyin Wang +6
The Diffusion Model (DM) has emerged as the SOTA approach for image synthesis. However, the existing DM cannot perform well on some image-to-image translation (I2I) tasks. Differen…
SOC: Semantic-Assisted Object Cluster for Referring Video Object Segmentation
Zhuoyan Luo, Yicheng Xiao, Yong Liu +5
This paper studies referring video object segmentation (RVOS) by boosting video-level visual-linguistic alignment. Recent approaches model the RVOS task as a sequence prediction pr…
FreeSeg: Unified, Universal and Open-Vocabulary Image Segmentation
Jie Qin, Jie Wu, Pengxiang Yan +8
Recently, open-vocabulary learning has emerged to accomplish segmentation for arbitrary categories of text-based descriptions, which popularizes the segmentation system to more gen…
DiffIR: Efficient Diffusion Model for Image Restoration
Bin Xia, Yulun Zhang, Shiyin Wang +5
Diffusion model (DM) has achieved SOTA performance by modeling the image synthesis process into a sequential application of a denoising network. However, different from image synth…
Global Knowledge Calibration for Fast Open-Vocabulary Segmentation
Kunyang Han, Yong Liu, Jun Hao Liew +8
Recent advancements in pre-trained vision-language models, such as CLIP, have enabled the segmentation of arbitrary concepts solely from textual inputs, a process commonly referred…
Global Spectral Filter Memory Network for Video Object Segmentation
Yong Liu, Ran Yu, Jiahao Wang +4
This paper studies semi-supervised video object segmentation through boosting intra-frame interaction. Recent memory network-based methods focus on exploiting inter-frame temporal…