18 citations · 62 across the 11 of their papers we have counts for
5 papers · 1 filter
OneFormer: One Transformer to Rule Universal Image Segmentation
Jitesh Jain, Jiachen Li, MangTik Chiu +3
Universal Image Segmentation is not a new concept. Past attempts to unify image segmentation in the last decades include scene parsing, panoptic segmentation, and, more recently, n…
Transformer Based Multi-Grained Features for Unsupervised Person Re-Identification
Jiachen Li, Menglin Wang, Xiaojin Gong
Multi-grained features extracted from convolutional neural networks (CNNs) have demonstrated their strong discrimination ability in supervised person re-identification (Re-ID) task…
VMFormer: End-to-End Video Matting with Transformer
Jiachen Li, Vidit Goel, Marianna Ohanyan +3
Video matting aims to predict the alpha mattes for each frame from a given input video sequence. Recent solutions to video matting have been dominated by deep convolutional neural…
Point-to-Box Network for Accurate Object Detection via Single Point Supervision
Pengfei Chen, Xuehui Yu, Xumeng Han +7
Object detection using single point supervision has received increasing attention over the years. However, the performance gap between point supervised object detection (PSOD) and…
Neighborhood Attention Transformer
Ali Hassani, Steven Walton, Jiachen Li +2
We present Neighborhood Attention (NA), the first efficient and scalable sliding-window attention mechanism for vision. NA is a pixel-wise operation, localizing self attention (SA)…