5 citations · 7 across the 3 of their papers we have counts for
3 papers
cs.CV2024
Multi-scale Contrastive Adaptor Learning for Segmenting Anything in Underperformed Scenes
Ke Zhou, Zhongwei Qiu, Dongmei Fu
Foundational vision models, such as the Segment Anything Model (SAM), have achieved significant breakthroughs through extensive pre-training on large-scale visual datasets. Despite…
cs.CV2023★ 5 cited
HAP: Structure-Aware Masked Image Modeling for Human-Centric Perception
Junkun Yuan, Xinyu Zhang, Hao Zhou +12
Model pre-training is essential in human-centric perception. In this paper, we first introduce masked image modeling (MIM) as a pre-training approach for this task. Upon revisiting…
cs.CV2023★ 2 cited
PSVT: End-to-End Multi-person 3D Pose and Shape Estimation with Progressive Video Transformers
Zhongwei Qiu, Yang Qiansheng, Jian Wang +6
Existing methods of multi-person video 3D human Pose and Shape Estimation (PSE) typically adopt a two-stage strategy, which first detects human instances in each frame and then per…