1 citations · 2 across the 5 of their papers we have counts for
5 papers
E4C: Enhance Editability for Text-Based Image Editing by Harnessing Efficient CLIP Guidance
Tianrui Huang, Pu Cao, Lu Yang +4
Diffusion-based image editing is a composite process of preserving the source image content and generating new content or applying modifications. While current editing approaches h…
CoT-MISR:Marrying Convolution and Transformer for Multi-Image Super-Resolution
Mingming Xiu, Yang Nie, Qing Song +1
As a method of image restoration, image super-resolution has been extensively studied at first. How to transform a low-resolution image to restore its high-resolution image informa…
Faster Learning of Temporal Action Proposal via Sparse Multilevel Boundary Generator
Qing Song, Yang Zhou, Mengjie Hu +1
Temporal action localization in videos presents significant challenges in the field of computer vision. While the boundary-sensitive method has been widely adopted, its limitations…
UV R-CNN: Stable and Efficient Dense Human Pose Estimation
Wenhe Jia, Yilin Zhou, Xuhan Zhu +3
Dense pose estimation is a dense 3D prediction task for instance-level human analysis, aiming to map human pixels from an RGB image to a 3D surface of the human body. Due to a larg…
SGM-Net: Semantic Guided Matting Net
Qing Song, Wenfeng Sun, Donghan Yang +2
Human matting refers to extracting human parts from natural images with high quality, including human detail information such as hair, glasses, hat, etc. This technology plays an e…