7 citations · 9 across the 2 of their papers we have counts for
3 papers
cs.CV2022★ 7 cited
Bridging CLIP and StyleGAN through Latent Alignment for Image Editing
Wanfeng Zheng, Qiang Li, Xiaoyan Guo +2
Text-driven image manipulation is developed since the vision-language model (CLIP) has been proposed. Previous work has adopted CLIP to design a text-image consistency-based object…
cs.CV2021
Improving Robustness for Pose Estimation via Stable Heatmap Regression
Yumeng Zhang, Li Chen, Yufeng Liu +3
Deep learning methods have achieved excellent performance in pose estimation, but the lack of robustness causes the keypoints to change drastically between similar images. In view…
cs.CV2020★ 2 cited
Improving Monocular Depth Estimation by Leveraging Structural Awareness and Complementary Datasets
Tian Chen, Shijie An, Yuan Zhang +4
Monocular depth estimation plays a crucial role in 3D recognition and understanding. One key limitation of existing approaches lies in their lack of structural information exploita…