4 citations · 4 across the 3 of their papers we have counts for
3 papers
cs.CV2023
AutoTaskFormer: Searching Vision Transformers for Multi-task Learning
Yang Liu, Shen Yan, Yuge Zhang +5
Vision Transformers have shown great performance in single tasks such as classification and segmentation. However, real-world problems are not isolated, which calls for vision tran…
cs.CV2023
Long-term Visual Localization with Mobile Sensors
Shen Yan, Yu Liu, Long Wang +6
Despite the remarkable advances in image matching and pose estimation, image-based localization of a camera in a temporally-varying outdoor environment is still a challenging probl…
cs.CV2022★ 4 cited
Multiview Transformers for Video Recognition
Shen Yan, Xuehan Xiong, Anurag Arnab +4
Video understanding requires reasoning at multiple spatiotemporal resolutions -- from short fine-grained motions to events taking place over longer durations. Although transformer…