2 citations · 3 across the 2 of their papers we have counts for
6 papers · 1 filter
E4C: Enhance Editability for Text-Based Image Editing by Harnessing Efficient CLIP Guidance
Tianrui Huang, Pu Cao, Lu Yang +4
Diffusion-based image editing is a composite process of preserving the source image content and generating new content or applying modifications. While current editing approaches h…
CoT-MISR:Marrying Convolution and Transformer for Multi-Image Super-Resolution
Mingming Xiu, Yang Nie, Qing Song +1
As a method of image restoration, image super-resolution has been extensively studied at first. How to transform a low-resolution image to restore its high-resolution image informa…
Faster Learning of Temporal Action Proposal via Sparse Multilevel Boundary Generator
Qing Song, Yang Zhou, Mengjie Hu +1
Temporal action localization in videos presents significant challenges in the field of computer vision. While the boundary-sensitive method has been widely adopted, its limitations…
UV R-CNN: Stable and Efficient Dense Human Pose Estimation
Wenhe Jia, Yilin Zhou, Xuhan Zhu +3
Dense pose estimation is a dense 3D prediction task for instance-level human analysis, aiming to map human pixels from an RGB image to a 3D surface of the human body. Due to a larg…
Renovating Parsing R-CNN for Accurate Multiple Human Parsing
Lu Yang, Qing Song, Zhihui Wang +5
Multiple human parsing aims to segment various human parts and associate each part with the corresponding instance simultaneously. This is a very challenging task due to the divers…
CPM R-CNN: Calibrating Point-guided Misalignment in Object Detection
Bin Zhu, Qing Song, Lu Yang +3
In object detection, offset-guided and point-guided regression dominate anchor-based and anchor-free method separately. Recently, point-guided approach is introduced to anchor-base…