201 citations · 215 across the 3 of their papers we have counts for
3 papers
cs.CV2023★ 1 cited
Going Denser with Open-Vocabulary Part Segmentation
Peize Sun, Shoufa Chen, Chenchen Zhu +4
Object detection has been expanded from a limited number of categories to open vocabulary. Moving forward, a complete intelligent vision system requires understanding more fine-gra…
cs.CV2021★ 13 cited
SLIP: Self-supervision meets Language-Image Pre-training
Norman Mu, Alexander Kirillov, David Wagner +1
Recent work has shown that self-supervised pre-training leads to improvements over supervised learning on challenging visual recognition tasks. CLIP, an exciting new approach to le…
cs.CV2021★ 201 cited
Masked Autoencoders Are Scalable Vision Learners
Kaiming He, Xinlei Chen, Saining Xie +3
This paper shows that masked autoencoders (MAE) are scalable self-supervised learners for computer vision. Our MAE approach is simple: we mask random patches of the input image and…