20 citations · 47 across the 10 of their papers we have counts for
Showing 2023 · cs.CVShow all
3 papers · 2 filters
cs.CV2023★ 2 cited
3D-Aware Neural Body Fitting for Occlusion Robust 3D Human Pose Estimation
Yi Zhang, Pengliang Ji, Angtian Wang +3
Regression-based methods for 3D human pose estimation directly predict the 3D pose parameters from a 2D image using deep networks. While achieving state-of-the-art performance on s…
cs.CV2023★ 1 cited
Cross-Modal Concept Learning and Inference for Vision-Language Models
Yi Zhang, Ce Zhang, Yushun Tang +1
Large-scale pre-trained Vision-Language Models (VLMs), such as CLIP, establish the correlation between texts and images, achieving remarkable success on various downstream tasks wi…
cs.CV2023★ 1 cited
From Association to Generation: Text-only Captioning by Unsupervised Cross-modal Mapping
Junyang Wang, Ming Yan, Yi Zhang +1
With the development of Vision-Language Pre-training Models (VLPMs) represented by CLIP and ALIGN, significant breakthroughs have been achieved for association-based visual tasks s…