31 citations · 39 across the 3 of their papers we have counts for
3 papers
cs.CV2022★ 8 cited
VMFormer: End-to-End Video Matting with Transformer
Jiachen Li, Vidit Goel, Marianna Ohanyan +3
Video matting aims to predict the alpha mattes for each frame from a given input video sequence. Recent solutions to video matting have been dominated by deep convolutional neural…
cs.CV2022
Point-to-Box Network for Accurate Object Detection via Single Point Supervision
Pengfei Chen, Xuehui Yu, Xumeng Han +7
Object detection using single point supervision has received increasing attention over the years. However, the performance gap between point supervised object detection (PSOD) and…
cs.CV2021★ 31 cited
SeMask: Semantically Masked Transformers for Semantic Segmentation
Jitesh Jain, Anukriti Singh, Nikita Orlov +4
Finetuning a pretrained backbone in the encoder part of an image transformer network has been the traditional approach for the semantic segmentation task. However, such an approach…