70 citations · 78 across the 3 of their papers we have counts for
3 papers
cs.CV2023★ 2 cited
Towards Robust Video Instance Segmentation with Temporal-Aware Transformer
Zhenghao Zhang, Fangtao Shao, Zuozhuo Dai +1
Most existing transformer based video instance segmentation methods extract per frame features independently, hence it is challenging to solve the appearance deformation problem. I…
cs.CV2022★ 6 cited
RenderNet: Visual Relocalization Using Virtual Viewpoints in Large-Scale Indoor Environments
Jiahui Zhang, Shitao Tang, Kejie Qiu +6
Visual relocalization has been a widely discussed problem in 3D vision: given a pre-constructed 3D visual map, the 6 DoF (Degrees-of-Freedom) pose of a query image is estimated. Re…
cs.CV2022★ 70 cited
QuadTree Attention for Vision Transformers
Shitao Tang, Jiahui Zhang, Siyu Zhu +1
Transformers have been successful in many vision tasks, thanks to their capability of capturing long-range dependency. However, their quadratic computational complexity poses a maj…