10 citations · 12 across the 2 of their papers we have counts for
3 papers
cs.CV2022★ 10 cited
SideRT: A Real-time Pure Transformer Architecture for Single Image Depth Estimation
Chang Shu, Ziming Chen, Lei Chen +3
Since context modeling is critical for estimating depth from a single image, researchers put tremendous effort into obtaining global context. Many global manipulations are designed…
cs.CV2021
Twins: Revisiting the Design of Spatial Attention in Vision Transformers
Xiangxiang Chu, Zhi Tian, Yuqing Wang +5
Very recently, a variety of vision transformer architectures for dense prediction tasks have been proposed and they show that the design of spatial attention is critical to their s…
cs.CV2021★ 2 cited
SwiftNet: Real-time Video Object Segmentation
Haochen Wang, Xiaolong Jiang, Haibing Ren +2
In this work we present SwiftNet for real-time semisupervised video object segmentation (one-shot VOS), which reports 77.8% J &F and 70 FPS on DAVIS 2017 validation dataset, leadin…