activity
20172021
most citedTFPose: Direct Human Pose Estimation with Transformers

57 citations · 203 across the 8 of their papers we have counts for

collaborators

9 papers

cs.CV20211 cited

SOLO: A Simple Framework for Instance Segmentation

Xinlong Wang, Rufeng Zhang, Chunhua Shen +2

Compared to many other dense prediction tasks, e.g., semantic segmentation, it is the arbitrary number of instances that has made instance segmentation much more challenging. In or…

cs.CV20212 cited

FCPose: Fully Convolutional Multi-Person Pose Estimation with Dynamic Instance-Aware Convolutions

Weian Mao, Zhi Tian, Xinlong Wang +1

We propose a fully convolutional multi-person pose estimation framework using dynamic instance-aware convolutions, termed FCPose. Different from existing methods, which often requi…

cs.CV202157 cited

TFPose: Direct Human Pose Estimation with Transformers

Weian Mao, Yongtao Ge, Chunhua Shen +3

We propose a human pose estimation framework that solves the task in the regression-based fashion. Unlike previous regression-based methods, which often fall behind those state-of-…

cs.CV20204 cited

Diverse Knowledge Distillation for End-to-End Person Search

Xinyu Zhang, Xinlong Wang, Jia-Wang Bian +2

Person search aims to localize and identify a specific person from a gallery of images. Recent methods can be categorized into two groups, i.e., two-step and end-to-end approaches.…

cs.CV202028 cited

BoxInst: High-Performance Instance Segmentation with Box Annotations

Zhi Tian, Chunhua Shen, Xinlong Wang +1

We present a high-performance method that can achieve mask-level instance segmentation with only bounding-box annotations for training. While this setting has been studied in the l…

cs.CV2020

End-to-End Video Instance Segmentation with Transformers

Yuqing Wang, Zhaoliang Xu, Xinlong Wang +4

Video instance segmentation (VIS) is the task that requires simultaneously classifying, segmenting and tracking object instances of interest in video. Recent methods typically deve…